Science sandboxes measure the scientific capability of AI agents
Arya S. Rao, Rodrigo I. Castro, Sager J. Gosai, Kenneth B. Hsu, Yasha Ektefaie, Shantanu Singh, Sangeeta N. Bhatia, Steven K. Reilly, Ryan Tewhey, Eric S. Lander, Pardis C. Sabeti
Abstract
Scientific progress depends not only on finding solutions, but on learning the rules that explain why they work and using that understanding to design better experiments. We introduce science sandboxes, a framework for studying this capability in AI agents through repeated cycles of experimentation, feedback, and hypothesis revision. Science sandboxes invite an agent to query the natural world in different ways, ranging from "wet" physical experiments, to "damp" predictive models trained on empirical data, to "dry" invented rules. By establishing a common experimental loop and a protocol for evaluating agents within it, science sandboxes allow assessment of both quantitative performance on specific metrics and qualitative scientific reasoning, across a spectrum of empirical verifiability. Here, we instantiate this framework in two biological settings, models of regulatory genomics and protein fitness prediction, and examine the capabilities of frontier agents. Across these settings, we could see when agents successfully optimized a quantitative metric without understanding the rules underlying the system. In particular, their scientific reasoning deteriorated when they encountered systems whose rules fell outside familiar biological priors. By highlighting such failure modes, science sandboxes make the frontier of scientific capability measurable and provide a controlled setting in which to study and ultimately expand it.
Create a lesson
Related papers
CryoAnomaly: Few-Shot Cryo-EM Particle Picking via Anomaly-Guided Hard Negative Suppression
Riku Itsuji, Rintaro Otsubo, Ryo Fujii et al.
FoldKit: A Python library for efficient storage and retrieval of co-folding predictions
Jonathan A. Levine, Melissa Pathil, Samuel Nitz et al.
Efficient Auto-Interpretability of AI Models in Biology
Piotr Jedryszek, Oliver M. Crook
CIR-DDG: backbone-agnostic residual correction of antibody-antigen affinity changes with explicit cross-chain geometry
Weilun Yu, Zhiheng Zou, Yonggui Huang et al.
Surf2Volume: a workflow for converting CIFTI parcellations to NIfTI volume space
Shuguang Yang, Ziyi Wang, Yujing Shen et al.
DINIRS: Digital Twin for Individualized Treatment Effects of Non-Invasive Respiratory Support Strategies
Md Fantacher Islam, Jarrod Mosier, Vignesh Subbian