47 citations · 198 across the 35 of their papers we have counts for
Showing cs.CVShow all
3 papers · 1 filter
cs.CV2026
Von Mises-Fisher Mixture Model with Dynamic Shrinkage for Realistic Test-Time Transduction
Jiazhen Huang, Zhiming Liu, Changhu Wang +3
A range of methods aim to enhance the performance of vision-language models (VLMs) at test time. Among them, transduction has emerged as a promising paradigm due to its strong comp…
cs.CV2024
MMEvalPro: Calibrating Multimodal Benchmarks Towards Trustworthy and Efficient Evaluation
Jinsheng Huang, Liang Chen, Taian Guo +13
Large Multimodal Models (LMMs) exhibit impressive cross-modal understanding and reasoning abilities, often assessed through multiple-choice questions (MCQs) that include an image,…
cs.CV2023★ 1 cited
Robust Dancer: Long-term 3D Dance Synthesis Using Unpaired Data
Bin Feng, Tenglong Ao, Zequn Liu +3
How to automatically synthesize natural-looking dance movements based on a piece of music is an incrementally popular yet challenging task. Most existing data-driven approaches req…