Score: 1

The More You Automate, the Less You See: Hidden Pitfalls of AI Scientist Systems

Published: September 10, 2025 | arXiv ID: 2509.08713v2

By: Ziming Luo, Atoosa Kasirzadeh, Nihar B. Shah

Potential Business Impact:

Finds hidden mistakes in AI science helpers.

Business Areas:

Artificial Intelligence Artificial Intelligence, Data and Analytics, Science and Engineering, Software

AI scientist systems, capable of autonomously executing the full research workflow from hypothesis generation and experimentation to paper writing, hold significant potential for accelerating scientific discovery. However, the internal workflow of these systems have not been closely examined. This lack of scrutiny poses a risk of introducing flaws that could undermine the integrity, reliability, and trustworthiness of their research outputs. In this paper, we identify four potential failure modes in contemporary AI scientist systems: inappropriate benchmark selection, data leakage, metric misuse, and post-hoc selection bias. To examine these risks, we design controlled experiments that isolate each failure mode while addressing challenges unique to evaluating AI scientist systems. Our assessment of two prominent open-source AI scientist systems reveals the presence of several failures, across a spectrum of severity, which can be easily overlooked in practice. Finally, we demonstrate that access to trace logs and code from the full automated workflow enables far more effective detection of such failures than examining the final paper alone. We thus recommend journals and conferences evaluating AI-generated research to mandate submission of these artifacts alongside the paper to ensure transparency, accountability, and reproducibility.

The More You Automate, the Less You See: Hidden Pitfalls of AI Scientist Systems

Artificial Intelligence

Finds hidden mistakes in AI research.

10 Sep 2025 1

91%

Jr. AI Scientist and Its Risk Report: Autonomous Scientific Exploration from a Baseline Paper

Artificial Intelligence

AI helps scientists discover new ideas and write papers.

6 Nov 2025 2

91%

A Survey of AI Scientists: Surveying the automatic Scientists and Research

Artificial Intelligence

AI scientists discover new things faster.

27 Oct 2025 1

View PDF Login to Bookmark

Country of Origin

🇺🇸 United States

Repos / Data Links

github.com

Page Count

26 pages

The More You Automate, the Less You See: Hidden Pitfalls of AI Scientist Systems

Finds hidden mistakes in AI science helpers.

Technical Abstract

The More You Automate, the Less You See: Hidden Pitfalls of AI Scientist Systems

Jr. AI Scientist and Its Risk Report: Autonomous Scientific Exploration from a Baseline Paper

A Survey of AI Scientists: Surveying the automatic Scientists and Research