2 papers
cs.RO2026
Fast and Accurate: An Adaptive VLA Inference Framework through Environment-aware Model Selection
Yuewei Sun, Lang Qin, Zechuan Tian +11
Embodied intelligence demands both long-horizon reasoning and real-time closed-loop responsiveness. Recent dual-system Vision-Language-Action (VLA) architectures combine fast react…
cs.AI2026
AllMem: A Memory-centric Recipe for Efficient Long-context Modeling
Ziming Wang, Xiang Wang, Kailong Peng +5
Large Language Models (LLMs) encounter significant performance bottlenecks in long-sequence tasks due to the computational complexity and memory overhead inherent in the self-atten…