Skip to content
AI Atlas
PaperActive

Sci-MMR: Benchmarking Multi-Step Evidence-Grounded Scientific Reasoning in Multimodal Agents

arxiv.org/abs/2609.11243

quality89

Updated 2 h ago · first seen 12 Sept 2026

paper_01M29X34J849J6ZBQPGG8SCEDM

Published
12 Sept 2026
T1 · 2 h ago
arXiv
2609.11243
T1 · 2 h ago
Category
cs.AI
T1 · 2 h ago

As of

Rewind the record: see this entity's attributes exactly as AI Atlas knew them on a given day.

Claim history · Authors

1 claims · 1 propertiesShow all properties

Authorsauthors1

Claim history for Authors
ValueValid from → toStatusSourceConfidenceExtractor
Bicheng Deng, Dingwei Zhu, Enyu Zhou, Han Wang, Jiadong Chen, Jiaqiang Li, Jiazheng Zhang, Lei Bai, Qi Zhang, Senjie Jin, Tao Gui, Xiang Zheng, Xingjun Ma, Yajie Yang, Yang Nan, Yanxin Li, Yuhui Wang, Zhiheng XicurrentcurrentarXiv (Atom API + RSS)T1highdeterministic

Claims are temporal and append-only: a new observation closes the previous claim (valid_to) instead of overwriting it. Conflicting claims from different sources are kept side by side and flagged — never averaged. Methodology →