arXiv:2606. 26448v1 Announce Type: cross Abstract: Across the sciences, autonomous systems are increasingly being used in closed-loop discovery, proposing new theories and designing and running experiments to test them.
By Akshay K. Jagadish, Younes Strittmatter, Nori Jacoby, George Kachergis, Eric Schulz, Nathaniel Daw, Suyog H. Chandramouli, Thomas L. Griffiths
arXiv:2606. 26460v1 Announce Type: new Abstract: AI-based scientific automation is increasingly possible by using agents to generate hypotheses, design experiments, and analyze data.
By Ben Prystawski, Kushin Mukherjee, Daniel Wurgaft, Linas Nasvytis, Michael Y. Li, Noah D. Goodman, Michael C. Frank
arXiv:2608. 07542v1 Announce Type: new Abstract: Autonomous research loops driven by large language models can run machine-learning experiments at scale but tend to drift toward local refinements of whichever metric they optimise rather than testing the hypotheses that motivate the experiments.
By Yiwen Zhang, Eloise Zeng, Jaeha Lee, Tony Yue Yu
arXiv:2604.27927v2 Announce Type: replace
Abstract: We introduce a framework called LAPITHS (Language model Analysis through Paradigm grounded Interpretations of Theses about Human likenesS) and use...
By Matteo Da Pelo, Alessio Donvito, Claudio Frongia, Pietro Salis, Antonio Lieto
arXiv:2510. 02660v2 Announce Type: replace-cross Abstract: When researchers claim AI systems possess ToM or mental models, they are fundamentally discussing behavioral predictions and bias corrections rather than genuine mental states.
By Xiaoyun Yin, Elmira Zahmat Doost, Shiwen Zhou, Garima Arya Yadav, Jamie C. Gorman
arXiv:2511. 04500v3 Announce Type: replace Abstract: Large language models (LLMs) are increasingly deployed as decision-making agents in high-stakes domains and as imitators of human behavior in the social and behavioral sciences.
By Andrea Cera Palatsi, Samuel Martin-Gutierrez, Ana S. Cardenal, Max Pellert
arXiv:2607. 26179v1 Announce Type: cross Abstract: LLMs are widely regarded as alien intelligences, systems whose cognitive operations are fundamentally unlike our own.
By Chandra Sripada, Richard Lewis
The article discusses how artificial intelligence is beginning to automate scientific discovery, specifically in the realm of cognitive science. It outlines four key challenges for developing an automated science of the mind: representing experiments, generating synthetic behavior, synthesizing models, and closing the loop to discover psychological theories. The authors argue that addressing these challenges will enable AI to systematically advance our understanding of the mind.
By Akshay K. Jagadish, Milena Rmus, Kristin Witte, Marvin Mathony, Marcel Binz, Eric Schulz
arXiv:2606. 11217v1 Announce Type: cross Abstract: The proliferation of large language models (LLMs) and autonomous AI agents has given rise to a rapidly growing methodological paradigm: "in silico" behavioral experiments.
By Michelle Vaccaro
The paper introduces eXplainable DFT (XDFT), a self‑evolving computational agent that transforms experiment‑simulation mismatches into executable searches for physical mechanisms. XDFT formalizes candidate mechanisms as hypotheses, tests them against experimental data, and refines its search strategy through a learning loop. In a benchmark of 112 cases where standard calculations predicted a metal but experiments found a semiconductor, XDFT resolved 105 cases with evidence‑supported mechanisms, and its top‑ranked hypotheses improved dramatically over initial expert priors.
By Yue Li, Penghui Yang, Yushan Xiao, Zhonghan Zhang, Jianguo Huang, Yuhao Lu, Cuntai Guan, Bo An, Bijun Tang, Zheng Liu
arXiv:2506. 21571v3 Announce Type: replace-cross Abstract: Large Reasoning Models (LRMs), which autonomously produce a reasoning Chain of Thought (CoT) before producing final responses, offer a promising approach to interpreting and monitoring model behaviors.
By Jianshuo Dong, Yujia Fu, Chuanrui Hu, Chao Zhang, Han Qiu
arXiv:2602.00685v2 Announce Type: replace
Abstract: Large language models (LLMs) are increasingly used as simulated participants in social science experiments, but their behavior is often unstable an...
By Xuan Liu, Haoyang Shang, Zizhang Liu, Xinyan Liu, Yunze Xiao, Yiwen Tu, Haojian Jin