AI for science
AI as an instrument of scientific discovery — AlphaFold solving protein structure, Olympiad-level mathematics, and by 2026 original results including a conjecture counterexample and formally-verified proofs.
This thread follows AI used not as a product but as a scientific instrument. Its landmark is AlphaFold. DeepMind’s AlphaFold 2 solved protein-structure prediction at the CASP14 assessment to a degree much of the field had thought a decade away, and its publication in Nature with open code and a database made it a standard tool across biology within months.
The programme widened. AlphaFold 3 extended prediction across proteins, DNA, RNA and small molecules; AlphaProof and AlphaGeometry 2 reached silver-medal standard at the International Mathematical Olympiad. The scientific establishment ratified the shift when Demis Hassabis and John Jumper shared the 2024 Nobel Prize in Chemistry for the protein work.
By 2026 the claim had moved from prediction to discovery. Anthropic reported that its Fable model had produced a counterexample to the Jacobian Conjecture, and OpenAI published ten formally-verified mathematical advances from an unreleased model. The open question the thread raises is where assistance ends and authorship begins — whether these systems are powerful instruments in human hands or are starting to do the science themselves.
Google Health reports AI matching radiologists on breast screening
A Nature paper claimed an AI system reduced false positives and false negatives in mammography against expert readers.
Models & capabilities · Benchmarks & progress
DeepMind publishes early AlphaFold protein structure work
Describes the CASP13-winning system, which used deep networks trained on genomic data to predict inter-residue distances rather than folding a structure directly.
Models & capabilities
The CORD-19 research dataset is released
Assembled with the White House OSTP, National Library of Medicine and Chan Zuckerberg Initiative, it opened with over 29,000 papers, most with full text.
Ideas & essays · Open weights & ecosystem
AlphaFold 2 solves protein structure prediction at CASP14
DeepMind's system predicted protein structures to roughly experimental accuracy, ending a fifty-year-old open problem in biology.
Models & capabilities · Benchmarks & progress
AlphaFold 2 is published in Nature and open-sourced
The method behind DeepMind's CASP14 result seven months earlier was released in full, with source code, rather than kept as a demonstrated but undisclosed system.
Models & capabilities · Open weights & ecosystem
The AlphaFold Protein Structure Database opens
The initial release covered around 350,000 predicted structures across the human proteome and twenty other organisms, with a stated plan to expand to over 100 million.
Open weights & ecosystem
DeepMind controls a fusion plasma with reinforcement learning
A single network commanding all of a tokamak's control coils held plasma shapes on Switzerland's TCV reactor, including configurations conventional controllers struggle with.
Models & capabilities · Ideas & essays
AlphaFold's database expands to 200 million structures
The database grew roughly 200-fold in a single release, from about a million structures to predictions covering nearly every catalogued protein across around a million species.
Open weights & ecosystem
AlphaTensor discovers new matrix multiplication algorithms
Found a 76-multiplication algorithm for a specific matrix size, improving on the best known method for the first time in over fifty years.
Models & capabilities
Meta pulls Galactica after three days
Its errors were formatted exactly like real citations and papers, so only an expert reader could tell fabrication from fact — a distinct failure mode from earlier chatbots' obvious mistakes.
Models & capabilities · Culture & impact
AlphaDev discovers faster sorting algorithms, added to the C++ library
The new sequences were merged into LLVM's libc++ standard library, its first change to that section of code in over a decade and the first written by a reinforcement-learning system.
Models & capabilities
AlphaMissense catalogues 71 million genetic variants for disease risk
Adapted from AlphaFold, the model classified 89% of all possible human missense variants as likely pathogenic or benign, versus 0.1% confirmed by human experts.
Models & capabilities
GraphCast beats conventional weather forecasting on speed and accuracy
Trained on four decades of reanalysis data, the model beat the ECMWF's physics-based system on over 90% of tested variables while running on a single TPU.
Models & capabilities
GNoME finds 2.2 million candidate new materials
DeepMind flagged 380,000 of the 2.2 million predicted crystal structures as most stable; independent labs had already synthesised 736 by the time of publication.
Models & capabilities
AlphaGeometry solves olympiad geometry problems near gold-medal level
The system solved 25 of 30 benchmark problems within competition time limits, versus the 25.9 average for human gold medalists and 10 for the prior best system.
Models & capabilities · Benchmarks & progress
AlphaFold 3 predicts structures across proteins, DNA, RNA and ligands
Restricted at launch to a rate-limited web server rather than downloadable code, prompting an open letter with more than 650 signatures within a week.
Models & capabilities
AlphaProof and AlphaGeometry 2 reach silver-medal standard at the IMO
The systems scored 28 of 42 points, one short of gold, but took up to three days on some problems against the competition's 4.5-hour limit.
Models & capabilities · Benchmarks & progress
DeepMind's AlphaProteo designs novel protein binders
Trained on the Protein Data Bank and over 100 million AlphaFold-predicted structures, the system succeeded on a cancer-linked target, VEGF-A, where prior methods had failed entirely.
Models & capabilities
Hassabis and Jumper share the Nobel Prize in Chemistry
Half the prize went to Baker for computational protein design; the other half was split between Hassabis and Jumper for AlphaFold's structure prediction.
Culture & impact
DeepMind reduces quantum computing errors with AlphaQubit decoder
Trained on Google's 49-qubit Sycamore processor, the Transformer-based decoder cut errors 6% versus the most accurate prior method and 30% versus the fastest, but remains too slow for real-time use.
Models & capabilities
Google Research launches an AI co-scientist to help generate hypotheses
A multi-agent Gemini 2.0 system with Generation, Reflection and Ranking agents proposed drug candidates later confirmed active in laboratory tests for two diseases.
Models & capabilities
Isomorphic Labs raises $600 million to advance AI-designed drugs toward clinical trials
Thrive Capital led the round, Isomorphic's first from outside Alphabet, with proceeds earmarked for internal oncology and immunology programmes alongside partnered work with Eli Lilly and Novartis.
Money & business
DeepMind's AlphaEvolve pairs Gemini with automated evaluators to discover algorithms
Not released to the public; DeepMind said the system had already been running inside Google, recovering 0.7% of worldwide data-centre compute and cutting Gemini training time.
Models & capabilities
Google launches Weather Lab with an experimental AI cyclone model
The experimental model produces 50 possible storm-path outcomes roughly a week ahead and was developed with feedback from the US National Hurricane Center.
Models & capabilities
DeepMind launches AlphaGenome for predicting genome regulatory activity
The model reads DNA sequences up to a million base pairs and predicts effects on gene splicing and expression at single-nucleotide resolution; weights followed for non-commercial use in January 2026.
Models & capabilities
Google DeepMind marks five years of AlphaFold's impact on biology
DeepMind said the AlphaFold Protein Structure Database, launched with over 200 million predicted structures, had been cited in more than 35,000 papers and drawn users in over 190 countries.
Models & capabilities
DeepMind's Genie 3 generates navigable, real-time interactive worlds
The system renders explorable 720p scenes at 24fps from a text prompt, holding roughly a minute of visual memory, and was released only to a small research cohort.
Models & capabilities
Trump launches Genesis Mission to accelerate AI-driven science
The order gives the Department of Energy 270 days to demonstrate an initial platform, after first identifying 20 science challenges and federal compute and data assets.
Government & policy · Compute & infrastructure
OpenAI introduces FrontierScience benchmark
GPT-5.2 scored 77% on olympiad-style questions but 25% on open-ended research tasks, a gap OpenAI's own researchers said showed little improvement over GPT-5.
Benchmarks & progress
OpenAI launches Prism
Built on Crixet, a LaTeX platform OpenAI had quietly acquired, Prism is free for any ChatGPT account and handles citation management and sketch-to-LaTeX conversion.
Models & capabilities
Shanghai AI Laboratory open-sources Intern-S1-Pro, a 1-trillion-parameter scientific model
Only 22 billion of the model's 1 trillion parameters activate per query; Shanghai AI Lab said it reaches gold-medal level on Olympiad-style mathematical and logical reasoning.
Open weights & ecosystem · Models & capabilities
Google upgrades Gemini 3 Deep Think to V2
Google reported 48.4% on Humanity's Last Exam without tools, 84.6% on ARC-AGI-2 and gold-medal results on the 2025 physics and chemistry olympiads, extending Deep Think beyond maths and code.
Models & capabilities
Anthropic tests Claude on BioMysteryBench
On 23 questions its own expert panel could not solve, an unreleased preview model Anthropic called Mythos scored roughly 30%, against single digits for Claude Haiku 4.5.
Benchmarks & progress
OpenAI model credited with disproving the Erdős unit distance conjecture
An internal general-purpose reasoning model found an infinite family of constructions beating a bound mathematicians had assumed near-optimal since 1946.
Benchmarks & progress · Models & capabilities · Ideas & essays
Google DeepMind's Co-Scientist reaches Nature publication as a multi-agent research tool
The Nature paper reported six case studies, including drug candidates that blocked 91% of a liver-scarring response, generated by a Generate-Debate-Evolve agent pipeline built on Gemini.
Models & capabilities
Google DeepMind's AlphaProof Nexus solves nine open Erdős problems
The system paired a language model with the Lean proof checker so every step is machine-verified, and solved each problem for a few hundred dollars in inference cost.
Models & capabilities · Ideas & essays
Anthropic's Fable model produces counterexample to the Jacobian Conjecture
Harvard mathematician Levent Alpöge said Claude Fable 5 found the three-variable counterexample in an evening; it disproves the conjecture from three dimensions upward.
Benchmarks & progress · Models & capabilities · Ideas & essays
Anthropic reports Claude finding novel cryptographic weaknesses
Anthropic's Frontier Red Team reports Claude Mythos Preview found a previously unknown attack halving the key strength of post-quantum scheme HAWK, and a new attack on round-reduced AES.
Security & misuse · Benchmarks & progress
Epoch AI expands FrontierMath to 50 unsolved research problems
Unlike FrontierMath's original tiers, these problems have no known solution at all; three of the fifty have been solved by AI, including one by GPT-5.6 Sol.
Benchmarks & progress
OpenAI publishes ten formally-verified math advances from unreleased Astra model
OpenAI said generating all ten proofs cost about $2,000 in compute; mathematician Gary Marcus called the framing 'vastly oversold' relative to what the paper actually verified.
Benchmarks & progress · Models & capabilities · Ideas & essays
DeepMind's WeatherNext model improves cyclone forecasting accuracy
WeatherNext gives roughly an extra day of predictive accuracy on cyclone track, intensity and wind structure versus prior forecasting models, per a Nature paper.
Models & capabilities