NVDA 196.51 ▼4.99%GOOGL 326.56 ▲2.13%MSFT 389.10 ▲1.94%AMD 494.95 ▼5.17%INTC 91.67 ▼0.70%TSMC 399.09 ▼1.07%AMZN 231.39 ▼0.31%META 593.87 ▼0.22%AAPL 336.91 ▲1.17%PLTR 131.53 ▲7.00%
Markets at last close

Google · Research

AI agents put pressure on science’s validation bottleneck

·1 min read

AI agents are moving from coding into scientific work, where they can plan tasks, call databases and tools, run subagents, write code and check their own errors. Google DeepMind’s Co-Scientist gave Imperial College London microbiologist José Penadés five possible explanations for how superbugs spread antibiotic resistance. Within two days, it matched the hypothesis his team had spent years developing: that some superbugs acquire viral tails and use them as keys to jump between host species.

Stronger frontier models, inference-time reasoning, agent scaffolding and custom skills are making these systems more useful to researchers. Agents can sift literature, query databases, automate analyses, generate hypotheses and search large solution spaces, as AlphaEvolve has done for TPU chip design, Erdős problems and genomics analysis. But LLM-based systems remain fallible, and researchers need transparency, confidence estimates and what Vivek Natarajan calls epistemic humility.

The central constraint is validation. Agents can make conjectures cheap and abundant, while wet lab experiments, peer review and facility access remain slow and costly. Policy priorities include broad access to agents, agent-ready data, investment in experimental infrastructure and automated labs, and reviewer tools that disclose AI use, cite evidence and preserve human judgment.

Originally reported by goo.gleRead the source →
Related coverage
All Google news →