3 papers
cs.CL2026
Random Attention: Rethinking KV Cache Eviction for Efficient Reasoning
Heng Wang, Jielin Qiu, Wenting Zhao +7
Large language models achieve superior performance on tasks that require extended reasoning, but long chains of thought make the KV cache a severe memory bottleneck. Existing KV ca…
cs.LG2026
AI4AI at Test-Time: Strong-to-Weak Capability Transfer via Harnesses
Cheng Qian, Wenting Zhao, Liangwei Yang +6
Recent work on distillation transfers the capabilities of large models to smaller ones often by updating the latter's parameters, through teacher forcing, on-policy distillation, a…
cs.RO2026
Mimir: A Neuro-Symbolic Memory System with Dynamic Grounding for Embodied Agents in Interactive Environments
Haoming Xu, Zhenlin He, Hengyi Wang +2
Long-horizon embodied task requires agents to act under partial observability while preserving both scene belief and execution progress. Flat histories or implicit policy states ma…