collaborators

12 papers

cs.LG2026

ResearchClawBench: A Benchmark for End-to-End Autonomous Scientific Research

Wanghan Xu, Shuo Li, Tianlin Ye +48

AI coding agents are increasingly used for scientific work, but their end-to-end autonomous research capability remains difficult to verify. We present ResearchClawBench, a benchma…

cs.AI2026

SCI-PRM: A Tool Aware Process Reward Model for Scientific Reasoning Verification

Xiangyu Zhao, Henry Hengyuan Zhao, Yiheng Wang +7

While Process Reward Models (PRMs) have achieved remarkable success in mathematical reasoning, their application in complex scientific domains-such as biology, chemistry, and physi…

hep-ph2026

Exponentially improved quantum simulation of scalar QFT

Qing-Hong Cao, Ying-Ying Li, Xiaohui Liu +2

Quantum simulations of scalar quantum field theories (QFT) provide important benchmarks for demonstrating quantum advantage. We revisit digitization in the occupation basis, which…

cs.CV2026

Omni-Weather: A Unified Multimodal Model for Weather Radar Understanding and Generation

Zhiwang Zhou, Yuandong Pu, Xuming He +10

Weather modeling requires both accurate prediction and mechanistic interpretation, yet existing methods treat these goals in isolation, separating generation from understanding. To…

cs.CL2026

PRBench: End-to-end Paper Reproduction in Physics Research

Shi Qiu, Junyi Deng, Yiwei Deng +48

AI agents powered by large language models exhibit strong reasoning and problem-solving capabilities, enabling them to assist scientific research tasks such as formula derivation a…

hep-ph2026

An End-to-end Architecture for Collider Physics and Beyond

Shi Qiu, Zeyu Cai, Jiashen Wei +7

We present, to our knowledge, the first language-driven agent system capable of executing end-to-end collider phenomenology tasks, instantiated within a decoupled, domain-agnostic…