2 papers
cs.CL2026
EndPrompt: Efficient Long-Context Extension via Terminal Anchoring
Han Tian, Luxuan Chen, Xinran Chen +10
Extending the context window of large language models typically requires training on sequences at the target length, incurring quadratic memory and computational costs that make lo…
cs.AI2026
No Action Without a NOD: A Heterogeneous Multi-Agent Architecture for Reliable Service Agents
Zixu Yang, Hang Zheng, Nan Jiang +5
Large language model (LLM) agents have increasingly advanced service applications, such as booking flight tickets. However, these service agents suffer from unreliability in long-h…