3d reconstruction 1edge acceleration 1hardware for AR/VR 1multi-precision compute 1transformer quantization 1
From the 1 of 9 linked papers with an AI index.
Showing cs.CVShow all
2 papers · 1 filter
cs.CV2026
Not All Tasks Quantize Equally: Fisher-Guided Quantization for Visual Geometry Transformer
Yipu Zhang, Jintao Cheng, Weilun Feng +5
Feed-forward 3D reconstruction models, represented by Visual Geometry Grounded Transformer (VGGT), jointly predict multiple visual geometry tasks such as depth estimation, camera p…
cs.CV2026
Training-Free Interaction-Aligned Visual Token Pruning for Efficient Embodied Manipulation
Jintao Cheng, Haozhe Wang, Weibin Li +7
Efficient visual representation is a central image-processing challenge in embodied manipulation, where policies repeatedly process dense visual-token sequences during closed-loop…