What is this?
AI for science means models that speed up discovery itself. They propose hypotheses, prove theorems, predict weather, find new materials and read whole literatures in minutes. Many researchers believe this, not chatbots, is where AI changes the world most.
Key tools & players
- DeepMind — AlphaFold, GNoME (materials), GraphCast (weather), AlphaProof (math)
- OpenAI — frontier models for research; a program giving 100,000 academics free access (our coverage)
- Anthropic — Claude for scientific reasoning and security research
- FutureHouse and "AI co-scientist" projects — autonomous research agents
Milestones
- 2020-2023 — AlphaFold2 opens the era, then GNoME and GraphCast take on materials and weather
- 2024-2025 — AlphaProof reaches Olympiad silver, then LLMs take IMO gold and co-scientist agents deliver lab-validated hypotheses
- 2026 — Claude finds encryption weaknesses experts missed, and OpenAI opens frontier models to 100k researchers (our coverage)
- Aug-Sep 2026 — AI takes the prime-gap record, FrontierMath's hardest tier and a machine-checked proof of Fermat (our coverage)
- Aug 2026 — First complete AI-designed genomes: 16 working bacteriophages, published in Science (our coverage)
- Aug-Sep 2026 — DeepMind's WeatherNext goes operational at the US Hurricane Center, then WeatherNext 3 delivers hourly 5 km global forecasts (our coverage)
- Sep 2026 — Claude finishes a particle-physics calculation one layer past the record a SLAC physicist held (our coverage)
- Sep 2026 — Claude turns 30-year-old Magellan radar into a public elevation map of a third of Venus
- Sep 2026 — 25 Fields medallists protest as OpenAI answers with an advisory group and 100 solved problems (our coverage)
- Sep 2026 — Claude rewrites 30+ biology models in a month, then finds a new phage enzyme system in Anthropic's own lab (our coverage)
Mini glossary
- Co-scientist: an AI agent that proposes and helps test hypotheses
- Foundation model: one large model adapted to many scientific tasks
- Benchmark: standardized test used to compare model abilities