gpt-5.4-nano-2026-03-17| Provider | Input, $ per 1M | Output, $ per 1M | In our data since |
|---|---|---|---|
| 302.AI | $0.2 | $1.25 | 31 Jul 2026 |
| Microsoft Foundry | $0.2 | $1.25 | 31 Jul 2026 |
| OpenAI | $0.2 | $1.25 | 31 Jul 2026 |
Only the base tier and only per-token prices are shown. Batch, discounted and cached rates, as well as prices per image or per second of video, do not go into this table: they cannot stand in the same column as a price per million tokens.
The date in the last column is the day this price first entered our collection. The price may well be older: before that day we simply were not recording it. It has not changed since — otherwise a new row with a new date would stand in its place.
| Task set | Result | Run conditions | Measured by |
|---|---|---|---|
| Mock AIME 2024–2025 — olympiad problems | 87.77 % | effort: high · with a tuned harness | Epoch evaluations |
| GPQA Diamond — graduate-level questions | 71.29 % | effort: high · with a tuned harness | Epoch evaluations |
| ARC-AGI — generalising to unseen patterns | 51.5 % | effort: xhigh · with a tuned harness | https://arcprize.org/leaderboard |
| WeirdML — unusual machine learning tasks | 49.23 % | effort: high | https://htihle.github.io/weirdml.html |
| FrontierMath, levels 1–3 | 44.91 % | effort: high | Epoch evaluations |
| Chess puzzles | 26.35 % | effort: high | Epoch evaluations |
| APEX-Agents | 16.9 % | — | — |
| FrontierMath, level 4 — research-grade problems | 12.2 % | effort: high | Epoch evaluations |
| SimpleQA Verified — factual accuracy | 12 % | effort: high | Epoch evaluations |
| CritPt — physics problems | 9.25 % | effort: xhigh | — |
| ARC-AGI-2 | 5.69 % | effort: xhigh | — |