NVDA 218.36 ▼2.37%GOOGL 332.60 ▲0.59%MSFT 492.44 ▲0.16%AMD 503.60 ▼3.36%INTC 100.32 ▼5.57%TSMC 428.03 ▼1.68%AMZN 251.89 ▼0.20%META 644.38 ▼1.42%AAPL 326.57 ▲3.56%PLTR 165.86 ▼2.16%
Markets at last close

Google · Research

Historical AI models struggle with Einstein-style discovery

·1 min read

Scientists are testing whether language models trained only on historical knowledge can rediscover major scientific breakthroughs. Demis Hassabis of Google DeepMind suggested using data available before 1911 to see whether an LLM could reproduce general relativity, framing the exercise as a possible benchmark for artificial general intelligence.

Early efforts suggest current systems are better at pattern matching than at the abductive reasoning behind paradigm shifts. Researchers argue that breakthroughs such as relativity require a creative leap from sparse or anomalous evidence into a durable world model, not just statistical extrapolation from large data sets. An MIT orbital-mechanics model trained on synthetic planetary systems failed to infer Newtonian gravitation, instead producing a different incorrect law for each system.

Independent researcher Michael Hla trained Machina Mirabilis on pre-1900 data and reported only limited “glimpses of intuition” after prompts related to quantum mechanics and relativity. Other teams found that historical training sets are noisy and leaky: a model intended to use only knowledge up to 1930 still answered later questions correctly. The Ranke-4B project has built models with cut-offs of 1913, 1929, 1933, 1939 and 1946, aiming to detect “sparks of genius” rather than full breakthroughs.

Originally reported by nature.comRead the source →
Related coverage
All Google news →