Model
gemini-2.5-pro
Gemini 2.5 Pro was Google DeepMind's "thinking" model from March 2025, built to reason through intermediate steps before answering rather than replying directly. Its release marked Google's strongest competitive position of the period: the company reported that it topped the crowdsourced LMArena leaderboard on debut and led several reasoning and science benchmarks, after two years largely trailing OpenAI and Anthropic on public rankings. It also carried a one-million-token context window, and was later succeeded as the reasoning race moved on.
Appears alongside
Featured in threads
Tracks
- Models & capabilities 5
- Benchmarks & progress 4
- Open weights & ecosystem 1
- Money & business 1
Google releases Gemini CLI, an open-source terminal AI agent
Free personal accounts get 60 requests a minute and 1,000 a day against Gemini 2.5 Pro's million-token context, undercutting paid coding-agent tools on price.
Open weights & ecosystem · Models & capabilities
ARC Prize compares reasoning models with no clear winner
ARC-AGI-2 remained unsolved by every system tested, and which model looked best depended entirely on whether accuracy or cost per task was prioritised.
Benchmarks & progress
Google updates Gemini 2.5 Pro preview with improved coding performance
The update, internally labelled 06-05, also led coding benchmarks including Aider Polyglot and performed strongly on Humanity's Last Exam.
Models & capabilities · Benchmarks & progress
Google I/O puts Gemini into search and ships Veo 3
AI Mode rolled out to all US Search users, and Veo 3 became the first widely-used video model to generate synchronised dialogue and sound effects alongside the picture.
Models & capabilities
Google launches AI Ultra subscription plan at Google I/O
At $249.99 a month — twelve times the existing Google AI Pro tier — the plan bundled early Veo 3 access, 30TB of storage and YouTube Premium.
Money & business · Models & capabilities
ETH Zurich's 'Proof or Bluff?' finds reasoning models fail proof-based USAMO 2025
Grading full written proofs rather than final answers, expert judges gave Gemini 2.5 Pro 24% and every other tested model under 5%, out of a possible 100%.
Benchmarks & progress
Gemini 2.5 Pro takes the lead on reasoning benchmarks
Google's thinking model topped LMArena and several reasoning evaluations, its strongest competitive position of the period.
Models & capabilities · Benchmarks & progress