3 citations · 3 across the 2 of their papers we have counts for
2 papers
cs.CV2024
MoPE-CLIP: Structured Pruning for Efficient Vision-Language Models with Module-wise Pruning Error Metric
Haokun Lin, Haoli Bai, Zhili Liu +5
Vision-language pre-trained models have achieved impressive performance on various downstream tasks. However, their large model sizes hinder their utilization on platforms with lim…
cs.CV2023★ 3 cited
Diverse 3D Hand Gesture Prediction from Body Dynamics by Bilateral Hand Disentanglement
Xingqun Qi, Chen Liu, Muyi Sun +3
Predicting natural and diverse 3D hand gestures from the upper body dynamics is a practical yet challenging task in virtual avatar creation. Previous works usually overlook the asy…