2 papers
cs.MM2026
EditEmoTalk: Controllable Speech-Driven 3D Facial Animation with Continuous Expression Editing
Diqiong Jiang, Kai Zhu, Dan Song +3
Speech-driven 3D facial animation aims to generate realistic and expressive facial motions directly from audio. While recent methods achieve high-quality lip synchronization, they…
cs.RO2025
AerialMind: Towards Referring Multi-Object Tracking in UAV Scenarios
Chenglizhao Chen, Shaofeng Liang, Runwei Guan +6
Referring Multi-Object Tracking (RMOT) aims to achieve precise object detection and tracking through natural language instructions, representing a fundamental capability for intell…