DeepSeek V4 Flash
DeepSeek V4 Flash is DeepSeek's reasoning-capable model: $0.14/$0.28 per 1M input/output tokens, 1M token context window. It supports native tool calling, extended reasoning, open weights. Released 2026-04-24.
Input Price
$0.14
Output Price
$0.28
Cache Read Price
$0
Context Window
1M
All metrics
| Metric | Value | Confidence | Verified | Provenance receipt |
|---|---|---|---|---|
| Input Price | $0.14 | Single source | Jun 10, 2026 | Receipt |
| Output Price | $0.28 | Single source | Jun 10, 2026 | Receipt |
| Cache Read Price | $0 | Verified · 2+ sources | Jun 10, 2026 | Receipt |
| Context Window | 1M | Verified · 2+ sources | Jun 10, 2026 | Receipt |
| Max Output Tokens | 384K | Verified · 2+ sources | Jun 10, 2026 | Receipt |
| Tool Calling | Yes | Verified · 2+ sources | Jun 10, 2026 | Receipt |
| Extended Reasoning | Yes | Verified · 2+ sources | Jun 10, 2026 | Receipt |
Pricing
Pay-as-you-go API
$0.14per 1M input tokens
- Output
- $0.28 / 1M tokens
- Cache read
- $0 / 1M tokens
Data sources
Facts on this page are sourced from the following verified sources.
- APImodels.devLast synced 3d ago
Compare DeepSeek V4 Flash
Related models
#7Gemini 3.5 Flash
Gemini 3.5 Flash is Google's reasoning-capable model: $1.50/$9 per 1M input/output tokens, 1.0M token context window.
#9Claude Haiku 4.5 (latest)
Anthropic
Claude Haiku 4.5 (latest) is Anthropic's reasoning-capable model: $1/$5 per 1M input/output tokens, 200K token context window, 39.45% on SWE-Bench Pro.
#16Llama 4 Maverick 17B Instruct
Meta
Llama 4 Maverick 17B Instruct is Meta's general-purpose model: 1M token context window, 5.24% on SWE-Bench Pro.
#17o4-mini
OpenAI
o4-mini is OpenAI's reasoning-capable model: $1.10/$4.40 per 1M input/output tokens, 200K token context window.
Spotted an error? Suggest a correction.