DeepSeek V4 Flash
By DeepSeek · Visit provider →
DeepSeek V4 Flash offers frontier-level performance at an unprecedented price point. With cache-hit pricing as low as $0.0028 per 1M input tokens, it's one of the most cost-effective high-quality models available.
Key Strengths
- Exceptional cost efficiency
- 1M token context window
- Strong coding and reasoning
- Cache-hit pricing dramatically reduces costs
Best Use Cases
Pricing Details
| Metric | Value |
|---|---|
| Input (per 1M tokens) | $0.14 |
| Output (per 1M tokens) | $0.28 |
| Context window | 1M tokens |
| Provider | DeepSeek |
| Architecture | Proprietary MoE architecture, 1M context, optimized for extreme cost efficiency with context caching |
Cache hit: $0.0028 input
Benchmark Performance
Source: Artificial Analysis (independent benchmarking). Updated 2026-08-08.
Recent AI News
Google slashes Gemini 3.7 Flash pricing by 50% to $0.75 per 1M tokens
MarkTechPost · Aug 13, 2026
Google Ships Gemini 3.7 Flash Three Weeks After 3.6 — Coding Scores Jump 9 Points
Ars Technica AI · Aug 13, 2026
DeepSeek V4 Pro update doubles agent scores, then doubles some API prices
The Decoder · Aug 13, 2026
Grok 4.6 ties GPT-5.6 Sol at 61 points while costing 60% less
The Decoder · Aug 12, 2026
Nvidia's $28B Nemotron 4 bet: One trillion parameters to chase Moonshot and DeepSeek
The Decoder · Aug 12, 2026
Get smarter about AI
The sharpest AI news, curated daily. Delivered free to your inbox.
Pricing data is sourced from official API documentation and updated weekly. Actual costs may vary based on prompt caching, batch processing, service tiers, and regional pricing. Inblix is not affiliated with DeepSeek.