OpenAI

gpt-5-nano-2025-08-07

Made by OpenAI. Released 2025-08-07. Accessibility not recorded. We hold no price for this model: no seller we can read lists it any more, so what follows is what was measured, not what it cost.

What it scores, by house

BenchmarkHouseScoreEvaluatedScaffold
chess puzzlesEpoch AI0.3±0.042026-08-07high
chess puzzlesEpoch AI0.1±0.042026-07-13low
chess puzzlesEpoch AI0.0±0.012026-07-13minimal
frontiermathEpoch AI0.1±0.022025-10-30high
frontiermathEpoch AI0.1±0.022025-08-07medium
frontiermath tier 4Epoch AI0.0±0.002025-10-30high
frontiermath tier 4Epoch AI0.0±0.022025-08-07medium
frontiermath tier 4 v2Epoch AI0.0±0.022026-06-12high
frontiermath tiers 1 3 v2Epoch AI0.1±0.012026-08-27low
frontiermath tiers 1 3 v2Epoch AI0.0±0.012026-08-27minimal
frontiermath tiers 1 3 v2Epoch AI0.2±0.022026-06-12high
gpqa diamondEpoch AI0.5±0.042026-07-13minimal
gpqa diamondEpoch AI0.6±0.042026-07-13low
gpqa diamondEpoch AI0.7±0.032025-10-30high
gpqa diamondEpoch AI0.7±0.032025-08-07medium
math level 5near its ceilingEpoch AI0.9±0.002025-08-20high
math level 5near its ceilingEpoch AI1.0±0.002025-08-20medium
mystery game puzzlesEpoch AI0.1±0.022026-08-27minimal
mystery game puzzlesEpoch AI0.1±0.032026-08-27medium
mystery game puzzlesEpoch AI0.1±0.032026-08-27high
mystery game puzzlesEpoch AI0.1±0.032026-08-27low
otis mock aime 2024 2025near its ceilingEpoch AI0.4±0.072026-07-13minimal
otis mock aime 2024 2025near its ceilingEpoch AI0.5±0.082026-07-13low
otis mock aime 2024 2025near its ceilingEpoch AI0.8±0.052025-10-31high
otis mock aime 2024 2025near its ceilingEpoch AI0.7±0.052025-08-07medium
simpleqa verifiedEpoch AI0.1±0.012026-08-10high
ale benchcollected by Epoch AI719high
arc aginear its ceilingcollected by Epoch AI0.2medium
arc aginear its ceilingcollected by Epoch AI0.0low
arc aginear its ceilingcollected by Epoch AI0.0minimal
arc aginear its ceilingcollected by Epoch AI0.2high
arc agi 2collected by Epoch AI0.0medium
arc agi 2collected by Epoch AI0.0high
arc agi 2collected by Epoch AI0.0minimal
arc agi 2collected by Epoch AI0.0low
dtbenchnear its ceilingcollected by Epoch AI0.6high
fictionlivebenchcollected by Epoch AI0.2medium
forecastbenchcollected by Epoch AI59.1unknown
lmcacollected by Epoch AI7.9high
proofbenchcollected by Epoch AI0.1high
terminalbenchcollected by Epoch AI0.1OpenHands
terminalbenchcollected by Epoch AI0.1Terminus 2
terminalbenchcollected by Epoch AI0.1Mini-SWE-Agent
terminalbenchcollected by Epoch AI0.2spoox-o-m
terminalbenchcollected by Epoch AI0.1Codex CLI
vpctcollected by Epoch AI0.4medium
vpctcollected by Epoch AI0.4high
weirdmlcollected by Epoch AI0.4high
weirdmlcollected by Epoch AI0.3low

These scores are not comparable between houses and are not added up. A score is dated by when it was evaluated, not when the model was released, and a score without its scaffold is not a measurement — which is why both are printed.

What retires, and when

  • deprecation2026-12-11, replaced by gpt-5.6-luna

Every rate we hold · Why there is no single best model ·