
DeepSeek V4 Flash Vision Exp is an experimental vision-enabled version of DeepSeek V4 Flash 0731Opens in new tab from DeepSeek, adding image understanding while matching the base model on text capabilities including agents, reasoning, and world knowledge. It is a sparse mixture-of-experts model with 13B active parameters out of 284B total.
It is suited for document and chart understanding, visual question answering, and multimodal agent workflows that interleave text and images.
| $0.44$0.2156 | $1.32$0.6468 | $0.014$0.00686 | 1.41s | 69 tps | ||
| $0.44 | $1.32 | $0.014 | 4.56s | 54 tps | ||
| $0.44 | $1.32 | $0.028 | 1.43s | 114 tps | ||
| $0.44 | $1.32 | $0.028 | 2.17s | 40 tps |
P50, best across providers
P50, best provider
When an error occurs in an upstream provider, we can recover by routing to another healthy provider, if your request filters allow it. You can access per-provider uptime data programmatically through the Endpoints API. Learn more about our load balancing and customization options.
