activity
20242026
collaborators

5 papers

cs.CV2026

Goodbye Drift: Anchored Tree Sampling for Long-Horizon Video-to-Video Generation

Matthew Bendel, Stephen W. Bailey, Mithilesh Vaidya +2

Long-horizon video generation suffers from two intertwined issues. First, there is drift, where video quality degrades over time. Second, there are continuity issues which manifest…

eess.AS2026

PoDAR: Power-Disentangled Audio Representation for Generative Modeling

Alejandro Luebs, Mithilesh Vaidya, Ishaan Kumar +5

The performance of audio latent diffusion models is primarily governed by generator expressivity and the modelability of the underlying latent space. While recent research has focu…

cs.CV2025

DreamTexture: Shape from Virtual Texture with Analysis by Augmentation

Ananta R. Bhattarai, Xingzhe He, Alla Sheffer +1

DreamFusion established a new paradigm for unsupervised 3D reconstruction from virtual views by combining advances in generative models and differentiable rendering. However, the u…

cs.CV2024

A Data Perspective on Enhanced Identity Preservation for Diffusion Personalization

Xingzhe He, Zhiwen Cao, Nicholas Kolkin +4

Large text-to-image models have revolutionized the ability to generate imagery using natural language. However, particularly unique or personal visual concepts, such as pets and fu…

cs.CV2024

LatentKeypointGAN: Controlling Images via Latent Keypoints

Xingzhe He, Bastian Wandt, Helge Rhodin

Generative adversarial networks (GANs) have attained photo-realistic quality in image generation. However, how to best control the image content remains an open challenge. We intro…