3 papers
cs.CV2024
T-SVG: Text-Driven Stereoscopic Video Generation
Qiao Jin, Xiaodong Chen, Wu Liu +2
The advent of stereoscopic videos has opened new horizons in multimedia, particularly in extended reality (XR) and virtual reality (VR) applications, where immersive content captiv…
cs.MM2024
ChatVTG: Video Temporal Grounding via Chat with Video Dialogue Large Language Models
Mengxue Qu, Xiaodong Chen, Wu Liu +2
Video Temporal Grounding (VTG) aims to ground specific segments within an untrimmed video corresponding to the given natural language query. Existing VTG methods largely depend on…
cs.CV2024
Motion Capture from Inertial and Vision Sensors
Xiaodong Chen, Wu Liu, Qian Bao +4
Human motion capture is the foundation for many computer vision and graphics tasks. While industrial motion capture systems with complex camera arrays or expensive wearable sensors…