Updated 4 h ago · first seen 16 Sept 2026
paper_01M2PPQEYV4G0C3907YK9PXYXR
Abstract
Autonomous research systems are increasingly capable of executing long research workflows, yet automation alone does not ensure that the resulting process remains scientifically grounded. We introduce AutoResearch, a two-stage system that connects Idea Generation with Idea Execution to address both how research ideas are formed and how they are reliably established through experimentation. In Idea Generation, AutoResearch continuously integrates emerging research signals with accumulated domain knowledge, identifies transferable mechanistic insights, and uses multi-model generation and cross-review to produce grounded, testable research plans. In Idea Execution, coordinated agents decompose these plans into experiments, iteratively implement and diagnose them, and employ independent evidence-based review before accepting research conclusions. Across representative settings in cross-modal retrieval, systems optimization, and benchmark-driven machine learning, AutoResearch turns generated ideas into measurable progress, detects and corrects unreliable experimental results, and makes evidence-conditioned decisions to continue, revise, or terminate research directions. For example, on RSICD benchmark, an AutoResearch-generated idea improves mean Recall from 32.84 to 34.69, while recording only 5 audit-confirmed issue events compared with 11-27 for other autonomous research systems. These results demonstrate a research process in which meaningful insight is grounded before experimentation and conclusions are grounded before acceptance: Insight In, Hallucination Out.
Organizations
Organizations 0
No organization stated. arXiv metadata does not carry affiliations; an organization is linked only when a model card or lab page cites the paper.
Models
Models introduced or described 0
Inbound described_by relations from model cards and documentation.
No model links this paper yet
Datasets
Datasets used 0
No dataset relation recorded.
Benchmarks
Benchmarks used 0
No benchmark relation recorded.
Code
Repositories & frameworks 0
No repository linked.
Timeline
Timeline 3
AutoResearch: Insight In, Hallucination Out: published at changed from 2026-09-16T04:00:00+00:00 to 2026-09-18T04:00:00+00:00
Published16 Sept 2026→18 Sept 2026arxivAutoResearch: Insight In, Hallucination Out: authors changed from ["Haoyang Zhang", "Jiahao Li", "Junjie Wang", "Qumeng Sun… to ["Bruno", "Haoyang Zhang", "Jiahao Li", "Qumeng Sun", "Xi…
AuthorsHaoyang Zhang, Jiahao Li, Junjie Wang, Qumeng Sun, Xiang Liu, Xiao Zhang, Yiming Ren→Bruno, Haoyang Zhang, Jiahao Li, Qumeng Sun, Xiang Liu, Xiao ZhangarxivNew paper: AutoResearch: Insight In, Hallucination Out
arxiv
Sources
Sources 1
Tier 1 = official/primary, 2 = quality secondary, 3 = community, 4 = unverified. Every snapshot is archived; see all sources and the methodology.