Model

claude-mythos-preview

7 entries · April 2026 – July 2026

Claude Mythos Preview, disclosed in April 2026, was Anthropic's most capable model at the time for coding and agentic tasks, but the company said it would not release it publicly because of the cyberattack capability it demonstrated — Anthropic reported it wrote a working Firefox exploit in 181 of several hundred attempts, against two for its predecessor Opus 4.6. Rather than ship it, Anthropic ran it against partners' code through a defensive programme, Project Glasswing, in what was widely described as the first model withheld on capability grounds since GPT-2 in 2019.

Tracks

  • Security & misuse 4
  • Safety & alignment 4
  • Models & capabilities 4
  • Benchmarks & progress 2