o4-mini
OpenAI's efficient reasoning model — near-o3 quality at a fraction of the cost for coding and STEM.
Pricing
Input$1.10 / 1M checked Jul 19, 2026
Cached input$0.275 / 1M
Output$4.40 / 1M checked Jul 19, 2026
Reasoning token usage is included in output cost.
Cost bandBalanced
Free tierNo free API tier.
Limits
Context window200k checked Jul 19, 2026
Max output100k
Knowledge cutoffJun 1, 2024
DeploymentOpenAI API
OpenAI-compatible APIYes
ModalitiesText, Image input
Where it fits
Best forHigh-volume reasoning tasks, coding agents, and STEM workflows where o3 cost is prohibitive.
Business fitThe best value reasoning model in the OpenAI lineup for production workloads.
Capabilities
Extended thinkingFunction callingStructured outputsVision
Benchmark scores
Composite88 / 100
Coding91 / 100 (SWE-bench Verified)
Reasoning93 / 100 (GPQA Diamond)
Long context84 / 100 (RULER 200K)
Vision82 / 100 (MMMU)
Instruction following90 / 100 (IFEval)
Editorial aggregate of published benchmark results as of May 1, 2025. These are not independent measurements run by this site.
Change history
No recorded changes for this model yet. Pricing and context are checked daily; changes appear here with dates and sources.
Data
Last checkedJul 19, 2026
Cross-checked againstopenrouter.ai/openai/o4-mini
Verify with provider documentation before committing spend. Report an error and it gets fixed with a logged correction.