From the 1 of 17 linked papers with an AI index.
3 papers · 1 filter
DiFlowDubber: Discrete Flow Matching for Automated Video Dubbing via Cross-Modal Alignment and Synchronization
Ngoc-Son Nguyen, Thanh V. T. Tran, Jeongsoo Choi +3
Video dubbing requires content accuracy, expressive prosody, high-quality acoustics, and precise lip synchronization, yet existing approaches struggle on all four fronts. To addres…
TESGNN: Temporal Equivariant Scene Graph Neural Networks for Efficient and Robust Multi-View 3D Scene Understanding
Quang P. M. Pham, Khoi T. N. Nguyen, Lan C. Ngo +3
Scene graphs have proven to be highly effective for various scene understanding tasks due to their compact and explicit representation of relational information. However, current m…
ESGNN: Towards Equivariant Scene Graph Neural Network for 3D Scene Understanding
Quang P. M. Pham, Khoi T. N. Nguyen, Lan C. Ngo +2
Scene graphs have been proven to be useful for various scene understanding tasks due to their compact and explicit nature. However, existing approaches often neglect the importance…