Model

o1

9 entries · September 2024 – February 2025

o1, released in September 2024, was the first of OpenAI's "reasoning" models, trained with reinforcement learning to work through a long internal chain of thought before answering. OpenAI reported large gains on maths, coding and science problems and, more consequentially, showed that accuracy kept rising the longer the model was allowed to think — a second way to convert compute into capability, separate from making models bigger. OpenAI chose to hide the raw reasoning trace from users, which limited outside scrutiny of how it worked, and rival labs published their own reasoning models within months.

Tracks

  • Models & capabilities 5
  • Benchmarks & progress 3
  • Safety & alignment 3
  • Ideas & essays 1
  • Security & misuse 1
  • Money & business 1