2 papers
cs.AI2026
Hera: Learning Long-Horizon Coordination for Device-Cloud Collaborative LLM Agents
Yuxin Zhang, Mengxue Hu, Zheng Lin +8
Large language model (LLM) agents excel at solving complex long-horizon tasks through autonomous interaction with environments. However, their real-world deployment faces a fundame…
cs.CL2026
Learning an Efficient Multi-Turn Dialogue Evaluator from Multiple LLM Judges
Yuqi Tang, Kehua Feng, Yunfeng Wang +6
Evaluating the conversational abilities of large language models (LLMs) remains a challenging task. Current mainstream approaches primarily rely on the "LLM-as-a-judge" paradigm, w…