3 papers
cs.CV2026
FlexiAvatar: Unified 3D Gaussian Human Avatars Under Arbitrary Body Visibility
Yihalem Yimolal Tiruneh, Muhammad Salman Ali, Uyoung Jeong +5
Reconstructing animatable 3D human avatars from monocular video is a fundamental problem in computer vision with broad applications in AR/VR and digital content creation. Existing…
cs.CV2026
Multi-THuMBS: Multi-person Tracking of 3D Human Meshes Beyond Video Shots
Jeongwan On, Muhammad Salman Ali, Muneeb A. Khan +6
Tracking multi-person 3D human meshes from in-the-wild videos is a highly challenging problem due to complex interactions, frequent occlusions, and severe truncation inherent in un…
cs.CV2026
HandVQA: Diagnosing and Improving Fine-Grained Spatial Reasoning about Hands in Vision-Language Models
MD Khalequzzaman Chowdhury Sayem, Mubarrat Tajoar Chowdhury, Yihalem Yimolal Tiruneh +4
Understanding the fine-grained articulation of human hands is critical in high-stakes settings such as robot-assisted surgery, chip manufacturing, and AR/VR-based human-AI interact…