DeepSeekVerified Active: September 18, 2026

DeepSeek V4-FlashAPI Pricing & Specifications

Ultra-economical high-speed inference engine with 1M context and 384K output window.

📊 Core Specifications & Pricing Rates

Standard Input
$0.14 / 1M tokens
Standard Output
$0.28 / 1M tokens
Cached Input
$0.0028 / 1M tokens
Context Window
1M (1,000,000 tokens)
Max Output Tokens
384,000 tokens
Supported Capabilities:✗ Vision✓ Function Calling / Tools✓ JSON Mode✓ Deep Reasoning

Live Workload Cost Estimator: DeepSeek V4-Flash

Adjust token volumes to simulate your exact monthly API bill.

Interactive Simulation
Estimated 60% repeated context
Estimated Total Monthly Cost
$4.96/ month
$0.00020 per API request(Saves $3.09/mo with caching)
Compare All Models in Calculator

💡 Estimated Monthly Workload Scenarios

High-Volume Code Parsing

$42.00/mo

100,000 queries (2,000 input, 500 output)

100,000 requests • 2000 in / 500 out

⚖️ Head-to-Head Competitor Comparisons

DeepSeek V4-Flash vs. Gemini 3.1 Flash-Lite

DeepSeek V4-Flash is 44% cheaper on input and 81% cheaper on output with 384K max output capacity.

Disruptive sub-cent economicsView Gemini 3.1 Flash-Lite

📜 Historical Price Changes & Updates

2026-07-15
DeepSeek V4-Flash launched at $0.14 in / $0.28 out per 1M tokens.

🔗 Verified Primary Sources

All pricing, token limits, and capabilities are cross-checked against official provider documentation and live API endpoints.

DeepSeek Official API Pricing