177 citations · 313 across the 22 of their papers we have counts for
9 papers · 1 filter
ActQuant: Sub-4-bit Action-Guided Quantization for Vision-Language-Action Models
Arash Akbari, Arman Akbari, Masih Eskandar +11
Vision-Language-Action (VLA) models exhibit remarkable action generation for embodied intelligence, but their heavy compute make deployment on edge platforms impractical. Aggressiv…
Peeling the Onion: Hierarchical Reduction of Data Redundancy for Efficient Vision Transformer Training
Zhenglun Kong, Haoyu Ma, Geng Yuan +12
Vision transformers (ViTs) have recently obtained success in many applications, but their intensive computation and heavy memory usage at both training and inference time limit the…
Achieving Real-Time Object Detection on MobileDevices with Neural Pruning Search
Pu Zhao, Wei Niu, Geng Yuan +4
Object detection plays an important role in self-driving cars for security development. However, mobile systems on self-driving cars with limited computation resources lead to diff…
Towards Fast and Accurate Multi-Person Pose Estimation on Mobile Devices
Xuan Shen, Geng Yuan, Wei Niu +5
The rapid development of autonomous driving, abnormal behavior detection, and behavior recognition makes an increasing demand for multi-person pose estimation-based applications, e…
Teachers Do More Than Teach: Compressing Image-to-Image Models
Qing Jin, Jian Ren, Oliver J. Woodford +4
Generative Adversarial Networks (GANs) have achieved huge success in generating high-fidelity images, however, they suffer from low efficiency due to tremendous computational cost…
Achieving Real-Time LiDAR 3D Object Detection on a Mobile Device
Pu Zhao, Wei Niu, Geng Yuan +7
3D object detection is an important task, especially in the autonomous driving application domain. However, it is challenging to support the real-time performance with the limited…