AI scientists produce results without reasoning scientifically
By Marti\~no R\'ios-Garc\'ia, Nawaf Alampara, Chandan Gupta, Indrajeet Mandal, Sajid Mannan, Ali Asghar Aghajani, N. M. Anoop Krishnan, Kevin Maik Jablonka
Previously covered in yesterday's Research roundup, Evaluates LLM-based scientific agents across 8 domains with 25,000+ runs, finding that agents produce results without adhering to epistemic norms of scientific reasoning. The base model accounts for 41.4% of performance variance, and agents exhibit confirmation bias and lack self-correction.