2 papers
cs.CL2026
Not All Thoughts Need HBM: Semantics-Aware Memory Hierarchy for LLM Reasoning
Aojie Yuan, Tianqi Shen, Dajun Zhang
Reasoning LLMs produce thousands of chain-of-thought tokens whose KV cache must reside in scarce GPU HBM. The dominant response -- permanently evicting low-importance tokens -- is…
cs.AI2026
AcademiClaw: When Students Set Challenges for AI Agents
Junjie Yu, Pengrui Lu, Weiye Si +75
Benchmarks within the OpenClaw ecosystem have thus far evaluated exclusively assistant-level tasks, leaving the academic-level capabilities of OpenClaw largely unexamined. We intro…