DeepSeek
deepseek-v4-flash-vision-exp API Pricing
DeepSeek charges $0.15 per 1M input tokens and $0.60 per 1M output tokens across a 1M context window. Verified 2026-09-15 against the official DeepSeek pricing page.
Vision model: images converted to tokens by dimensions, billed as input; cache-miss off-peak rate shown; cache-hit input $0.003 off-peak; peak hours double all rates
Input / 1M tokens
$0.15
Output / 1M tokens
$0.60
Context window
1M
What it costs to run
| Workload | Tokens | Cost |
|---|---|---|
| Single short call | 2K in, 500 out | $0.00060 |
| Typical chat turn | 10K in, 1K out | $0.00210 |
| Long document pass | 100K in, 2K out | $0.016 |
| 1,000 chat turns | 10M in, 1M out | $2.10 |
Price history
Input price has moved -32% since tracking began on 2026-08-26.
| Effective | Input / 1M | Output / 1M |
|---|---|---|
| 2026-09-15 | $0.15 | $0.60 |
| 2026-08-26 | $0.22 | $0.66 |
How it compares
Ranked on the blended cost of a typical chat turn (10K in, 1K out), not the input rate alone.
| gemini-2.0-flash-lite | $0.00105 per turn | |
| gemini-2.5-flash-lite | $0.00140 per turn | |
| OpenAI | gpt-4.1-nano | $0.00140 per turn |
| Anthropic | claude-haiku-5.5 | $0.00150 per turn |
| OpenAI | gpt-6-luna | $0.00150 per turn |
| DeepSeek | deepseek-v4-flash-vision-exp | $0.00210 per turn |
| OpenAI | gpt-5.6-luna | $0.00320 per turn |
| OpenAI | gpt-5.4-nano | $0.00325 per turn |
| gemini-3.1-flash-lite | $0.00400 per turn | |
| OpenAI | babbage-002 | $0.00440 per turn |
| OpenAI | gpt-5-mini | $0.00450 per turn |
TechFastForward coverage of deepseek-v4-flash-vision-exp
What we wrote when this model, or the rest of its line, was news.
- DeepSeek V4.1 Flash Beats Anthropic on Coding Tasks2026-10-05
- DeepSeek V4 Pro Undercuts US AI Coding Cost 95% 20262026-06-05
- Qwen3.7 Max Surpasses Gemini Flash on AI Index 20262026-06-05
- DeepSeek V4 Cuts Frontier AI Coding Cost 75 Percent2026-06-02
- DeepSeek V4 Pro Doesn't Just Match Claude, It Does It at 86% Off the Sticker Price2026-05-09
- DeepSeek V4 Flash Just Killed the Inference Pricing Premium, and Nobody Noticed2026-05-08
Free JSON API
Every tracked model, including this one, is available as JSON. Free to use with attribution and a link back.
curl https://techfastforward.com/api/data/llm-pricingMaintained by Jordan Hale, TechFastForward. Prices are read daily from the official provider page and recorded only when they move. See every tracked model on the LLM API Pricing Tracker.