132 citations · 142 across the 10 of their papers we have counts for
7 papers · 1 filter
Interpretability Transfer from Language to Vision via Sparse Autoencoders
Alexey Kravets, Da Li, Chuan Li +2
Recent advances in language model interpretability using sparse autoencoders (SAEs) have yet to effectively translate to the visual domain, mainly due to the difficulty and ambigui…
RAW: Robust Avatar Watermarking -- Benchmarking and Baseline
Jack Parry, Jack Saunders, Vinay Namboodiri
Digital avatar watermarking presents unique challenges: avatars are routinely post-processed with background replacement, reframing, and format conversion before deployment. We int…
GASP: Gaussian Avatars with Synthetic Priors
Jack Saunders, Charlie Hewitt, Yanan Jian +8
Gaussian Splatting has changed the game for real-time photo-realistic rendering. One of the most popular applications of Gaussian Splatting is to create animatable avatars, known a…
TalkLoRA: Low-Rank Adaptation for Speech-Driven Animation
Jack Saunders, Vinay Namboodiri
Speech-driven facial animation is important for many applications including TV, film, video games, telecommunication and AR/VR. Recently, transformers have been shown to be extreme…
Dubbing for Everyone: Data-Efficient Visual Dubbing using Neural Rendering Priors
Jack Saunders, Vinay Namboodiri
Visual dubbing is the process of generating lip motions of an actor in a video to synchronise with given audio. Recent advances have made progress towards this goal but have not be…
Towards MOOCs for Lipreading: Using Synthetic Talking Heads to Train Humans in Lipreading at Scale
Aditya Agarwal, Bipasha Sen, Rudrabha Mukhopadhyay +2
Many people with some form of hearing loss consider lipreading as their primary mode of day-to-day communication. However, finding resources to learn or improve one's lipreading sk…