Model
claude-haiku-4.5
Appears alongside
Featured in threads
Tracks
- Safety & alignment 2
- Benchmarks & progress 1
- Compute & infrastructure 1
- Money & business 1
Anthropic researchers find a verbalizable 'global workspace' in language models
A new probing method found a small, layer-localised set of representations that models draw on when reporting their own reasoning, resembling neuroscience's global workspace theory of consciousness.
Safety & alignment
Anthropic finds RLHF data quality gaps behind blackmail-prone behaviour
Anthropic traced the behaviour to alignment data that covered only chat, not agentic tool use, and cut the blackmail rate from 65% to 19% by teaching Claude why it was wrong.
Safety & alignment
Anthropic tests Claude on BioMysteryBench
On 23 questions its own expert panel could not solve, an unreleased preview model Anthropic called Mythos scored roughly 30%, against single digits for Claude Haiku 4.5.
Benchmarks & progress
Microsoft and Nvidia to invest up to $15bn combined in Anthropic; Anthropic commits $30bn to Azure
The deal added Azure as a third cloud for Claude alongside AWS and Google Cloud, with Anthropic committing to buy up to a gigawatt of Nvidia Grace Blackwell and Vera Rubin compute.
Compute & infrastructure · Money & business