DeepSeek
deepseek-v4-flash API Pricing
DeepSeek charges $0.14 per 1M input tokens and $0.28 per 1M output tokens across a 1M context window. Verified 2026-06-11 against the official DeepSeek pricing page.
Cache-miss input rate; cache-hit input is $0.0028
Input / 1M tokens
$0.14
Output / 1M tokens
$0.28
Context window
1M
What it costs to run
| Workload | Tokens | Cost |
|---|---|---|
| Single short call | 2K in, 500 out | $0.00042 |
| Typical chat turn | 10K in, 1K out | $0.00168 |
| Long document pass | 100K in, 2K out | $0.015 |
| 1,000 chat turns | 10M in, 1M out | $1.68 |
Price history
No price move recorded for deepseek-v4-flash since tracking began on 2026-06-11. Every future change lands here with the before and after.
How it compares
Ranked on the blended cost of a typical chat turn (10K in, 1K out), not the input rate alone.
| gemini-2.5-flash-lite | $0.00140 per turn | |
| DeepSeek | deepseek-v4-flash | $0.00168 per turn |
| OpenAI | gpt-5.6-luna | $0.00320 per turn |
| OpenAI | gpt-5.4-nano | $0.00325 per turn |
| gemini-3.1-flash-lite | $0.00400 per turn | |
| DeepSeek | deepseek-v4-pro | $0.00522 per turn |
| gemini-2.5-flash | $0.00550 per turn |
Free JSON API
Every tracked model, including this one, is available as JSON. Free to use with attribution and a link back.
curl https://techfastforward.com/api/data/llm-pricingMaintained by Jordan Hale, TechFastForward. Prices are read daily from the official provider page and recorded only when they move. See every tracked model on the LLM API Pricing Tracker.