activity
20222024
most citedMutual Adaptive Reasoning for Monocular 3D Multi-Person Pose Estimation

3 citations · 8 across the 5 of their papers we have counts for

collaborators

5 papers

cs.CV20243 cited

LLaVA-SLT: Visual Language Tuning for Sign Language Translation

Han Liang, Chengyu Huang, Yuecheng Xu +6

In the realm of Sign Language Translation (SLT), reliance on costly gloss-annotated datasets has posed a significant barrier. Recent advancements in gloss-free SLT methods have sho…

cs.CV2024

The Language of Motion: Unifying Verbal and Non-verbal Language of 3D Human Motion

Changan Chen, Juze Zhang, Shrinidhi K. Lakshmikanth +5

Human communication is inherently multimodal, involving a combination of verbal and non-verbal cues such as speech, facial expressions, and body gestures. Modeling these behaviors…

cs.CV2024

HOI-M3:Capture Multiple Humans and Objects Interaction within Contextual Environment

Juze Zhang, Jingyan Zhang, Zining Song +6

Humans naturally interact with both others and the surrounding multiple objects, engaging in various social activities. However, recent advances in modeling human-object interactio…

cs.CV20232 cited

IKOL: Inverse kinematics optimization layer for 3D human pose and shape estimation via Gauss-Newton differentiation

Juze Zhang, Ye Shi, Yuexin Ma +3

This paper presents an inverse kinematic optimization layer (IKOL) for 3D human pose and shape estimation that leverages the strength of both optimization- and regression-based met…

cs.CV20223 cited

Mutual Adaptive Reasoning for Monocular 3D Multi-Person Pose Estimation

Juze Zhang, Jingya Wang, Ye Shi +3

Inter-person occlusion and depth ambiguity make estimating the 3D poses of monocular multiple persons as camera-centric coordinates a challenging problem. Typical top-down framewor…