Skip to content

Model detail

deepseek-v4-flash-vision-exp

Provided by DeepSeek
Pay-as-you-go

DeepSeek, adding image understanding while matching the base model on text capabilities including agents, reasoning, and world knowledge. It is a sparse mixture-of-experts model with 13B active parameters out of 284B total. It is suited for document and chart understanding, visual question answering, and multimodal agent workflows that interleave text and images.

Model specs

Context length
1M
Max output
384K
I/O modalities
Text / Image / File
Released
2026-08

02

API endpoints

One MixRoute gateway, OpenAI-compatible

  • OpenAI-compatible /v1/chat/completions POST

03

Pricing

Flat rate, unit: /1M Tokens

Input

$0.4400 /1M Tokens

Completion

$1.3200 /1M Tokens

Cache read

$0.0140 /1M Tokens

08

Token cost estimator

Live estimate from this page's pricing, not an actual bill

Share of the same prefix read repeatedly, up to 100%

One endpoint, a testable decision

Test this model and alternate routes with the same request format

Start from a real workload, then let quality, total cost, and failure conditions decide whether to send production traffic.