1 paper
Junxiao Xue, Xiaozhen Liu, Xuecheng Wu +2
Estimating spoken content from silent videos is crucial for applications in Assistive Technology (AT) and Augmented Reality (AR). However, accurately mapping lip movement sequences…