2 papers
cs.CV2024
GenCA: A Text-conditioned Generative Model for Realistic and Drivable Codec Avatars
Keqiang Sun, Amin Jourabloo, Riddhish Bhalodia +9
Photo-realistic and controllable 3D avatars are crucial for various applications such as virtual and mixed reality (VR/MR), telepresence, gaming, and film production. Traditional m…
eess.AS2020
Multimodal active speaker detection and virtual cinematography for video conferencing
Ross Cutler, Ramin Mehran, Sam Johnson +4
Active speaker detection (ASD) and virtual cinematography (VC) can significantly improve the remote user experience of a video conference by automatically panning, tilting and zoom…