Home  /  Models  /  OpenAI o3-mini
⚡ Quick Pricing Summary

OpenAI o3-mini by OpenAI costs $1.10 per 1 Million input tokens and $4.40 per 1 Million output tokens. It features a 200k context window and has a blended 3:1 production rate of $1.925/1M tokens.

OpenAI Reasoning / Deep Thinking

OpenAI o3-mini

Next-generation cost-efficient reasoning model with adjustable reasoning effort levels.

Calculate Spend in App →
Input Token Price
$1.10
Per 1,000,000 tokens
Output Token Price
$4.40
Per 1,000,000 tokens
Prompt Caching Rate
$0.550
Save up to 80% on cached inputs
Max Context Window
200k
Max output: 100k tokens

Verified Specifications & Benchmark Data

Developer / Provider OpenAI
Model Family o-Series
Blended 3:1 Rate (Production Benchmark) $1.925 / 1M tokens
Batch API Discount (24hr SLA) 50% off standard rate
Multimodal Vision ❌ Text Only
Function Calling / Structured Outputs ✅ Native Tool Calling
MMLU Benchmark Score 89.5%
Latency & Throughput Tier Fast Reasoning (~60 t/s)
Optimal Architecture & Use Cases High-throughput complex code synthesis and mathematical logic

Frequently Asked Questions about OpenAI o3-mini

How much does OpenAI o3-mini cost per 1M tokens?

OpenAI o3-mini pricing is set at $1.10 per 1 million input tokens and $4.40 per 1 million output tokens. For high-volume batch workloads, 24-hour batch queues provide a 50% off standard rate.

How much money does OpenAI o3-mini prompt caching save?

With prompt caching enabled, cached input tokens are discounted to $0.550/1M, saving 50% on repeated system prompts and document vectors.

What is the context window limit of OpenAI o3-mini?

OpenAI o3-mini has a maximum context window of 200k (200,000 tokens), supporting up to 100,000 completion tokens per response.

Direct Matchups with OpenAI o3-mini

See how OpenAI o3-mini compares against other leading frontier and open-weights models in cost and latency.