AI scientists produce results without reasoning scientifically
By Marti\~no R\'ios-Garc\'ia, Nawaf Alampara, Chandan Gupta, Indrajeet Mandal, Sajid Mannan, Ali Asghar Aghajani, N. M. Anoop Krishnan, Kevin Maik Jablonka
Evaluates LLM-based scientific agents across 8 domains with 25,000+ agent runs, finding that agents produce results without adhering to epistemic norms of scientific reasoning. The base model accounts for 41.4% of performance variance, dominating the scaffold.