3 papers
cs.CV2026
Not All Attention Heads Contribute to Critical Visual Token Selection: Head-Aware Pruning Matters More
Chaofang Ma, Lin Jiang, Carol Jingyi Li +4
Vision-Language Models (VLMs) have exhibited impressive performance across diverse visual scenarios. However, this success comes at the cost of explosive growth in visual tokens, w…
cs.AR2026
VersaQ-3D: Architecture Support for Visual Geometry Grounded Transformers via Versatile Quantization
Yipu Zhang, Jintao Cheng, Xingyu Liu +8
3D reconstruction and view synthesis are fundamental to AR/VR, robotics, and digital twins. The Visual Geometry Grounded Transformer (VGGT) enables strong feed-forward 3D reconstru…
cs.AR2025
AMD Versal Implementations of FAM and SSCA Estimators
Carol Jingyi Li, Ruilin Wu, Philip H. W. Leong
Cyclostationary analysis is widely used in signal processing, particularly in the analysis of human-made signals, and spectral correlation density (SCD) is often used to characteri…