Showing cs.CVShow all
2 papers · 1 filter
cs.CV2026
ActQuant: Sub-4-bit Action-Guided Quantization for Vision-Language-Action Models
Arash Akbari, Arman Akbari, Masih Eskandar +11
Vision-Language-Action (VLA) models exhibit remarkable action generation for embodied intelligence, but their heavy compute make deployment on edge platforms impractical. Aggressiv…
cs.CV2024
Exploring Token Pruning in Vision State Space Models
Zheng Zhan, Zhenglun Kong, Yifan Gong +8
State Space Models (SSMs) have the advantage of keeping linear computational complexity compared to attention modules in transformers, and have been applied to vision tasks as a ne…