collaborators

5 papers

cs.CL2026

How Do Agents Fail on AutoResearch: End-to-End Diagnostic Evaluation on 100 Real-World Frontier Research Tasks

Yanlin Fei, Nazhou Liu, Xinmiao Yu +7

AI has long assisted scientific research, but the rapid advance of LLMs and agentic scaffolds is reshaping the landscape; a single system can now carry whole-stage research from an…

cs.AI2026

Reconstruction: A Blind Benchmark for Recovering Research Ideas from Pre-Publication Bibliographies

Shaolong Chen, Yanlin Fei, Nazhou Liu +7

Can a language model recover the true research idea of a published paper when given only that paper's pre-publication bibliography? We introduce Reconstruction, a blind idea-recove…

physics.chem-ph2026

Reaction-Network-Level Discovery of Ammonia Synthesis Catalysts via Ten-Million-Scale Generative Exploration

Ruili Li, Rui Qi, Shuoqi Zhang +5

Catalyst discovery for ammonia synthesis is inherently a reaction-network challenge because catalytic performance is governed not by a single adsorbed intermediate, but by a surfac…

cs.LG2026

CLaaS: Continual learning as a service for sample efficient online learning

Kion Fallah, Silen Naihin, Barak Widawsky +1

Deployed large language model agents must adapt to distribution shift in dynamic environments. Ideally, adaptation can be performed from accumulated agent experiences and retain pr…

cs.CL2026

State-of-the-Art Arabic Language Modeling with Sparse MoE Fine-Tuning and Chain-of-Thought Distillation

Navan Preet Singh, Anurag Garikipati, Ahmed Abulkhair +6

This paper introduces Arabic-DeepSeek-R1, an application-driven open-source Arabic LLM that leverages a sparse MoE backbone to address the digital equity gap for under-represented…