OpenAIVerified Active: September 18, 2026

GPT-5.1 MiniAPI Pricing & Specifications

Budget-tier inference engine with 400K context and 128K max output.

📊 Core Specifications & Pricing Rates

Standard Input
$0.25 / 1M tokens
Standard Output
$2.00 / 1M tokens
Cached Input
$0.0250 / 1M tokens
Context Window
400K (400,000 tokens)
Max Output Tokens
128,000 tokens
Supported Capabilities:✓ Vision✓ Function Calling / Tools✓ JSON ModeStandard Inference
📦 Asynchronous Batch API Supported: Official provider documentation offers a 50% discount on non-streaming 24h batch requests.

Live Workload Cost Estimator: GPT-5.1 Mini

Adjust token volumes to simulate your exact monthly API bill.

Interactive Simulation
Estimated 60% repeated context
Estimated Total Monthly Cost
$24.31/ month
$0.00097 per API request(Saves $5.06/mo with caching)
Compare All Models in Calculator

💡 Estimated Monthly Workload Scenarios

High-Throughput Content Moderation

$45.00/mo

200,000 text evaluations (500 input, 50 output)

200,000 requests • 500 in / 50 out

⚖️ Head-to-Head Competitor Comparisons

GPT-5.1 Mini vs. Gemini 3.1 Flash-Lite

Both models are identical on input ($0.25/1M); Gemini 3.1 Flash-Lite is cheaper on output ($1.50 vs $2.00) with 1M context.

Similar budget tierView Gemini 3.1 Flash-Lite

📜 Historical Price Changes & Updates

2026-08-01
Maintained on price sheet as low-cost budget workhorse alongside newer flagships.

🔗 Verified Primary Sources

All pricing, token limits, and capabilities are cross-checked against official provider documentation and live API endpoints.

OpenAI Developer Platform Pricing