AI Models Recover Only 3–15% of Research Ideas, Study Finds
A new benchmark testing seven frontier AI models found they could reconstruct a scientific paper's core idea from its pre-publication bibliography alone only 3-15% of the time.
Read more →This project is suspended: no new articles or editions will be published. The archive stays available.
Every AI News story tagged with both Large Language Models and AI Research — the two topics side by side, updated as new articles publish.
5 articles
A new benchmark testing seven frontier AI models found they could reconstruct a scientific paper's core idea from its pre-publication bibliography alone only 3-15% of the time.
Read more →Thinking Machines Lab's second open-weight model needs a quarter of its predecessor's active parameters yet edges past it on reasoning and coding benchmarks.
Read more →
The Turing Test asks whether a machine's conversation can fool a human judge. Here's how Alan Turing's 1950 thought experiment works, and what a 2025 study found when it tested GPT-4.5.
Read more →AI world models go beyond language to build internal simulations of physical reality — predicting what happens next when an action is taken. Here is what they are, how they work, and why researchers think they are essential for the next generation of AI.
Read more →A large language model is the AI technology behind ChatGPT, Claude, Gemini, and most modern AI assistants — this guide explains how it works and why it matters.
Read more →