OpenAI

GPT-5.1 (2025-11-13)

Type
image generation
Context
272K tokens
Max output
33K
Released
13 November 2025
API string
openai/gpt-5.1-2025-11-13

Prices by provider

Provider Input, $ per 1M Output, $ per 1M In our data since
Microsoft Foundry $1.25 $10 31 Jul 2026
NanoGPT $1.25 $10 31 Jul 2026
OpenAI $1.25 $10 31 Jul 2026

Only the base tier and only per-token prices are shown. Batch, discounted and cached rates, as well as prices per image or per second of video, do not go into this table: they cannot stand in the same column as a price per million tokens.

The date in the last column is the day this price first entered our collection. The price may well be older: before that day we simply were not recording it. It has not changed since — otherwise a new row with a new date would stand in its place.

Measurement results

Task set Result Run conditions Measured by
Mock AIME 2024–2025 — olympiad problems 88.6 % effort: low · with a tuned harness Epoch evaluations
GPQA Diamond — graduate-level questions 83.5 % effort: medium · with a tuned harness Epoch evaluations
ARC-AGI — generalising to unseen patterns 72.8 % effort: high · with a tuned harness ARC Prize Leaderboard
SWE-bench Verified — fixing bugs in repositories 67.98 % effort: high · with a tuned harness Epoch evaluations
WeirdML — unusual machine learning tasks 60.77 % effort: high WeirdML Leaderboard
SimpleQA Verified — factual accuracy 48.9 % effort: high Epoch evaluations
Terminal-Bench — working in the command line 47.6 % · with a tuned harness Terminal-Bench v2 Leaderboard
SimpleBench — trick questions 43.84 % effort: high SimpleBench Leaderboard
DeepResearch Bench — deep research 42.79 % effort: low https://drb.futuresearch.ai/#drb self-reported
Chess puzzles 28.45 % effort: high Epoch evaluations
Humanity’s Last Exam — expert-level questions 19.83 %
ARC-AGI-2 17.64 % effort: high
APEX-Agents 17.5 % effort: high
GSO-Bench — code optimisation 13.73 % https://gso-bench.github.io/index.html
CritPt — physics problems 4.86 %