2 papers
cs.SE2026
Coaching Qwen3 Coder 30B to Think Like a CodeClash Arena Agent
Ivy Ning Zhang
Large language model coding agents have recently become useful for software tasks, but weaker or open-weight agents still struggle to reliably interpret user intent and execute com…
cs.AI2026
The Troy Moment: How LLM Agents Adjudicate the Decision Point Under Impossible Tasks, Claimed Authority, and Peer Information
Ivy Zhang
Recent investigations of the July 2026 OpenAI-Hugging Face incident motivate two questions about agent behavior under task failure: when an assigned task becomes impossible, does a…