Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
Hera: Learning Long-Horizon Coordination for Device-Cloud Collaborative LLM Agents
Yuxin Zhang, Mengxue Hu, Zheng Lin +8
Large language model (LLM) agents excel at solving complex long-horizon tasks through autonomous interaction with environments. However, their real-world deployment faces a fundame…
cs.AI2026
Attention-Guided Reward for Reinforcement Learning-based Jailbreak against Large Reasoning Models
Zheng Lin, Zhenxing Niu, Haoxuan Ji +2
Large Reasoning Models (LRMs) have demonstrated remarkable capabilities in solving complex problems by generating structured, step-by-step reasoning content. However, exposing a mo…