5 papers
Optimizing and Fine-tuning Large Language Model for Urban Renewal
Xi Wang, Xianyao Ling, Tom Zhang +5
This study aims to innovatively explore adaptive applications of large language models (LLM) in urban renewal. It also aims to improve its performance and text generation quality f…
Model-aware 3D Eye Gaze from Weak and Few-shot Supervisions
Nikola Popovic, Dimitrios Christodoulou, Danda Pani Paudel +2
The task of predicting 3D eye gaze from eye images can be performed either by (a) end-to-end learning for image-to-gaze mapping or by (b) fitting a 3D eye model onto images. The fo…
EFE: End-to-end Frame-to-Gaze Estimation
Haldun Balim, Seonwook Park, Xi Wang +2
Despite the recent development of learning-based gaze estimation methods, most methods require one or more eye or face region crops as inputs and produce a gaze direction vector as…
FONT: Flow-guided One-shot Talking Head Generation with Natural Head Motions
Jin Liu, Xi Wang, Xiaomeng Fu +4
One-shot talking head generation has received growing attention in recent years, with various creative and practical applications. An ideal natural and vivid generated talking head…
OPT: One-shot Pose-Controllable Talking Head Generation
Jin Liu, Xi Wang, Xiaomeng Fu +4
One-shot talking head generation produces lip-sync talking heads based on arbitrary audio and one source face. To guarantee the naturalness and realness, recent methods propose to…