Showing cs.CVShow all
2 papers · 1 filter
cs.CV2026
State Beyond Appearance: Diagnosing and Improving State Consistency in Dial-Based Measurement Reading
Yuanze Hu, Gen Li, Yuqin Lan +5
Multimodal large language models (MLLMs) have achieved impressive progress on general multimodal tasks, yet they remain brittle on dial-based measurement reading. In this paper, we…
cs.CV2025
FaceEditTalker: Controllable Talking Head Generation with Facial Attribute Editing
Guanwen Feng, Zhiyuan Ma, Yunan Li +3
Recent advances in audio-driven talking head generation have achieved impressive results in lip synchronization and emotional expression. However, they largely overlook the crucial…