DeepSeekVerified Active: September 18, 2026
DeepSeek V4-FlashAPI Pricing & Specifications
Ultra-economical high-speed inference engine with 1M context and 384K output window.
📊 Core Specifications & Pricing Rates
Standard Input
$0.14 / 1M tokens
Standard Output
$0.28 / 1M tokens
Cached Input
$0.0028 / 1M tokens
Context Window
1M (1,000,000 tokens)
Max Output Tokens
384,000 tokens
Supported Capabilities:✗ Vision✓ Function Calling / Tools✓ JSON Mode✓ Deep Reasoning
Live Workload Cost Estimator: DeepSeek V4-Flash
Adjust token volumes to simulate your exact monthly API bill.
Estimated 60% repeated context
Estimated Total Monthly Cost
Compare All Models in Calculator→$4.96/ month
$0.00020 per API request(Saves $3.09/mo with caching)
💡 Estimated Monthly Workload Scenarios
High-Volume Code Parsing
$42.00/mo100,000 queries (2,000 input, 500 output)
100,000 requests • 2000 in / 500 out
⚖️ Head-to-Head Competitor Comparisons
DeepSeek V4-Flash vs. Gemini 3.1 Flash-Lite
DeepSeek V4-Flash is 44% cheaper on input and 81% cheaper on output with 384K max output capacity.
Disruptive sub-cent economicsView Gemini 3.1 Flash-Lite →
📜 Historical Price Changes & Updates
2026-07-15
DeepSeek V4-Flash launched at $0.14 in / $0.28 out per 1M tokens.
🔗 Verified Primary Sources
All pricing, token limits, and capabilities are cross-checked against official provider documentation and live API endpoints.