OpenAIVerified Active: September 18, 2026
GPT-5.1 MiniAPI Pricing & Specifications
Budget-tier inference engine with 400K context and 128K max output.
📊 Core Specifications & Pricing Rates
Standard Input
$0.25 / 1M tokens
Standard Output
$2.00 / 1M tokens
Cached Input
$0.0250 / 1M tokens
Context Window
400K (400,000 tokens)
Max Output Tokens
128,000 tokens
Supported Capabilities:✓ Vision✓ Function Calling / Tools✓ JSON ModeStandard Inference
📦 Asynchronous Batch API Supported: Official provider documentation offers a 50% discount on non-streaming 24h batch requests.
Live Workload Cost Estimator: GPT-5.1 Mini
Adjust token volumes to simulate your exact monthly API bill.
Estimated 60% repeated context
Estimated Total Monthly Cost
Compare All Models in Calculator→$24.31/ month
$0.00097 per API request(Saves $5.06/mo with caching)
💡 Estimated Monthly Workload Scenarios
High-Throughput Content Moderation
$45.00/mo200,000 text evaluations (500 input, 50 output)
200,000 requests • 500 in / 50 out
⚖️ Head-to-Head Competitor Comparisons
GPT-5.1 Mini vs. Gemini 3.1 Flash-Lite
Both models are identical on input ($0.25/1M); Gemini 3.1 Flash-Lite is cheaper on output ($1.50 vs $2.00) with 1M context.
Similar budget tierView Gemini 3.1 Flash-Lite →
📜 Historical Price Changes & Updates
2026-08-01
Maintained on price sheet as low-cost budget workhorse alongside newer flagships.
🔗 Verified Primary Sources
All pricing, token limits, and capabilities are cross-checked against official provider documentation and live API endpoints.