Google DeepMind

Gemini 2.0 Flash Thinking 0121

Type
language model
Context
1,000K tokens
Max output
8K
Released
21 January 2025
API string
gemini-2.0-flash-thinking-exp-01-21

Prices by provider

Provider Input, $ per 1M Output, $ per 1M In our data since
NanoGPT $0.306 $1.003 31 Jul 2026

Only the base tier and only per-token prices are shown. Batch, discounted and cached rates, as well as prices per image or per second of video, do not go into this table: they cannot stand in the same column as a price per million tokens.

The date in the last column is the day this price first entered our collection. The price may well be older: before that day we simply were not recording it. It has not changed since — otherwise a new row with a new date would stand in its place.

Measurement results

Task set Result Run conditions Measured by
Creative writing (Lech Mazur’s evaluation) 73.8 % lechmazur/writing Github repository
Mock AIME 2024–2025 — olympiad problems 57.74 % · with a tuned harness Epoch evaluations
Fiction.LiveBench — holding a long context 52.8 % Fiction.live leaderboard
GPQA Diamond — graduate-level questions 42.76 % · with a tuned harness Epoch evaluations
Aider Polyglot — code edits in six languages 18.2 % · with a tuned harness Aider LLM Leaderboards
SimpleBench — trick questions 16.84 % SimpleBench Leaderboard
Humanity’s Last Exam — expert-level questions 1.85 %