7 citations · 7 across the 13 of their papers we have counts for
4 papers · 1 filter
Spark-to-Paper: End-to-End Research Paper Generation as a Composable Skill
Zhuoyang Qian, Biao Wu, Yiran Wang +6
Turning a research idea into a complete paper requires more than text generation: the system must retrieve literature, design and execute experiments, revise claims according to ev…
PaperJury: Due-Process Review for Bounded LaTeX Revision
Yiran Wang, Ruixuan An, Biao Wu +1
Pre-submission hardening of human-authored LaTeX computer science papers differs from drafting assistance because it requires adversarial whole-paper review, explicit no-fix outcom…
I-WebGenBench : Evaluating Interactivity in LLM-Generated Scientific Web Applications
Dasen Dai, Biao Wu, Meng Fang +2
Recent advances in visual language models have enabled autonomous agents for complex reasoning, tool use, and document understanding. However, existing document agents mainly trans…
InfoMosaic-Bench: Evaluating Multi-Source Information Seeking in Tool-Augmented Agents
Yaxin Du, Yuanshuo Zhang, Xiyuan Yang +10
Information seeking is a fundamental requirement for humans. However, existing LLM agents rely heavily on open-web search, which exposes two fundamental weaknesses: online content…