2 papers
cs.CL2026
DASH-KV: Accelerating Long-Context LLM Inference via Asymmetric KV Cache Hashing
Jinyu Guo, Zhihan Zhang, Jiehui Xie +7
The quadratic computational complexity of the standard attention mechanism constitutes a fundamental bottleneck for large language models in long-context inference. While existing…
cs.MA2026
Gated Coordination for Efficient Multi-Agent Collaboration in Minecraft Game
HuaDong Jian, Chenghao Li, Haoyu Wang +4
In long-horizon open-world multi-agent systems, existing methods often treat local anomalies as automatic triggers for communication. This default design introduces coordination no…