12 papers · 1 filter
Training-free Controllable Human Motion Generation under Heterogeneous Constraints
Xiaofei Hui, Bo Yan, Haoxuan Qu +2
Training-free controllable motion generation has attracted growing interest for enabling flexible constraint enforcement without constraint-specific training. However, existing tra…
TraMP-LLaMA: Generative Interpretability with Decoupled Instruction Tuning for Facial Expression Quality Assessment
Shuchao Duan, Alan Whone, Hossein Rahmani +2
Existing facial expression quality assessment (FEQA) methods typically produce only a severity score, without explicitly communicating the observable facial motion evidence that su…
ToolFG: Towards Well-Grounded Fine-Grained Image Classification
Yu Xue, Haoxuan Qu, Zhuoling Li +4
Fine-grained image classification (FGIC) has broad applications and has attracted significant research attention. In this paper, we explore a novel paradigm for solving FGIC by pro…
Translating Signals to Languages for sEMG-Based Activity Recognition
Ming Wang, Haoxuan Qu, Qiuhong Ke +3
Surface electromyography (sEMG) signal-based activity recognition has attracted increasing research attention in recent years. To develop accurate sEMG signal-based activity recogn…
MicroscopyMatching: Towards a Ready-to-use Framework for Microscopy Image Analysis in Diverse Conditions
Xiaofei Hui, Haoxuan Qu, Hossein Rahmani +3
Analyzing microscopy images to extract biological object properties (e.g., their morphological organization, temporal dynamics, and population density) is fundamental to various bi…
When Visual Privacy Protection Meets Multimodal Large Language Models
Xiaofei Hui, Qian Wu, Haoxuan Qu +3
The emergence of Multimodal Large Language Models (MLLMs) and the widespread usage of MLLM cloud services such as GPT-4V raised great concerns about privacy leakage in visual data.…