OpenAI

o4-mini-2025-04-16

Made by OpenAI. Released 2025-04-16. Accessibility not recorded. Cheapest standard rate we hold: $8 per million output tokens, at data sharing context.

What it costs

TierKindInOutCached inFrom
batchtext$2$8$0.50seller
batchtext · data sharing$1$4$0.25seller
standardtext$4$16$1seller
standardtext · data sharing$2$8$0.50seller

per million tokens, US dollars. * marks a rate from a single registry or two that disagree — the best available number rather than a confirmed one. Every other row is confirmed against a seller or a licensed index.

What it scores, by house

BenchmarkHouseScoreEvaluatedScaffold
chess puzzlesEpoch AI0.2±0.042026-08-07medium
chess puzzlesEpoch AI0.1±0.032026-07-11low
chess puzzlesEpoch AI0.3±0.042025-12-08high
frontiermathEpoch AI0.1±0.022025-11-16low
frontiermathEpoch AI0.2±0.022025-11-13medium
frontiermathEpoch AI0.2±0.032025-11-13high
frontiermath tier 4Epoch AI0.0±0.022025-08-07medium
frontiermath tier 4Epoch AI0.1±0.042025-07-01high
frontiermath tier 4 v2Epoch AI0.0±0.032026-06-11high
frontiermath tiers 1 3 v2Epoch AI0.2±0.022026-08-27low
frontiermath tiers 1 3 v2Epoch AI0.3±0.032026-08-27medium
frontiermath tiers 1 3 v2Epoch AI0.4±0.032026-06-11high
gpqa diamondEpoch AI0.8±0.032026-08-07medium
gpqa diamondEpoch AI0.8±0.032026-07-11low
gpqa diamondEpoch AI0.8±0.022025-04-16high
math level 5near its ceilingEpoch AI1.0±0.002025-04-16high
mystery game puzzlesEpoch AI0.1±0.022026-08-27high
otis mock aime 2024 2025near its ceilingEpoch AI0.7±0.072026-08-07medium
otis mock aime 2024 2025near its ceilingEpoch AI0.6±0.072026-07-13low
otis mock aime 2024 2025near its ceilingEpoch AI0.8±0.052025-04-16high
simpleqa verifiedEpoch AI0.2±0.012026-08-27low
simpleqa verifiedEpoch AI0.2±0.012026-08-27high
aider polyglotcollected by Epoch AI72.0—high
ale benchcollected by Epoch AI826—high
algotunecollected by Epoch AI1.7—high
arc aginear its ceilingcollected by Epoch AI0.4—medium
arc aginear its ceilingcollected by Epoch AI0.6—high
arc aginear its ceilingcollected by Epoch AI0.2—low
arc agi 2collected by Epoch AI0.1—high
arc agi 2collected by Epoch AI0.0—medium
arc agi 2collected by Epoch AI0.0—low
cad evalcollected by Epoch AI0.6—medium
critptcollected by Epoch AI0.0—high
dtbenchnear its ceilingcollected by Epoch AI0.8—high
enigma evalcollected by Epoch AI0.1—medium
enigma evalcollected by Epoch AI0.1—high
fictionlivebenchcollected by Epoch AI0.6—medium
forecastbenchcollected by Epoch AI61.8—unknown
gdpvalcollected by Epoch AI0.3—high
geobenchcollected by Epoch AI3651—medium
geobenchcollected by Epoch AI3717—high
gsocollected by Epoch AI0.0—OpenHands
hlecollected by Epoch AI0.1—medium
hlecollected by Epoch AI0.2—high
lech mazur writingcollected by Epoch AI7.5—medium
lmcacollected by Epoch AI26.5—high
metr time horizonscollected by Epoch AI76.5—medium
simplebenchcollected by Epoch AI0.4—high
vpctcollected by Epoch AI0.6—medium
weirdmlcollected by Epoch AI0.5—high

These scores are not comparable between houses and are not added up. A score is dated by when it was evaluated, not when the model was released, and a score without its scaffold is not a measurement — which is why both are printed.

What retires, and when

  • deprecation — 2026-10-23
  • deprecation — 2026-10-23

Every rate we hold · Why there is no single best model · Everything from OpenAI