Showing cs.CVShow all
3 papers · 1 filter
cs.CV2025
PhonemeFake: Redefining Deepfake Realism with Language-Driven Segmental Manipulation and Adaptive Bilevel Detection
Oguzhan Baser, Ahmet Ege Tanriverdi, Sriram Vishwanath +1
Deepfake (DF) attacks pose a growing threat as generative models become increasingly advanced. However, our study reveals that existing DF datasets fail to deceive human perception…
cs.CV2025
Learnings from Scaling Visual Tokenizers for Reconstruction and Generation
Philippe Hansen-Estruch, David Yan, Ching-Yao Chung +7
Visual tokenization via auto-encoding empowers state-of-the-art image and video generative models by compressing pixels into a latent space. Although scaling Transformer-based gene…
cs.CV2024
Unified Auto-Encoding with Masked Diffusion
Philippe Hansen-Estruch, Sriram Vishwanath, Amy Zhang +1
At the core of both successful generative and self-supervised representation learning models there is a reconstruction objective that incorporates some form of image corruption. Di…