Showing cs.CLShow all
2 papers · 1 filter
cs.CL2026
How Do Agents Fail on AutoResearch: End-to-End Diagnostic Evaluation on 100 Real-World Frontier Research Tasks
Yanlin Fei, Nazhou Liu, Xinmiao Yu +7
AI has long assisted scientific research, but the rapid advance of LLMs and agentic scaffolds is reshaping the landscape; a single system can now carry whole-stage research from an…
cs.CL2026
AdaDPO: Self-Adaptive Direct Preference Optimization with Balanced Gradient Updates
Shaolong Chen, Madalina Ciobanu, Qingqing Mao +1
DPO has become a widely adopted alternative to RLHF for aligning LLMs with human preferences, eliminating the need for a separate reward model or RL loop. Recent theoretical analys…