Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
MixReasoning: Switching Modes to Think
Haiquan Lu, Gongfan Fang, Xinyin Ma +2
Reasoning models enhance performance by tackling problems in a step-by-step manner, decomposing them into sub-problems and exploring long chains of thought before producing an answ…
cs.AI2026
On-Policy Self-Evolution via Failure Trajectories for Agentic Safety Alignment
Bo Yin, Qi Li, Xinchao Wang
Tool-using LLM agents fail through trajectories rather than only final responses, as they may execute unsafe tool calls, follow injected instructions, comply with harmful requests,…