4 papers · 1 filter
VidStyleODE: Disentangled Video Editing via StyleGAN and NeuralODEs
Moayed Haji Ali, Andrew Bond, Tolga Birdal +4
We propose , a spatiotemporally continuous disentangled eo representation based upon GAN and Neural-s. Effective t…
GANFusion: Feed-Forward Text-to-3D with Diffusion in GAN Space
Souhaib Attaiki, Paul Guerrero, Duygu Ceylan +2
We train a feed-forward text-to-3D diffusion generator for human characters using only single-view 2D data for supervision. Existing 3D generative models cannot yet match the fidel…
SuperGaussian: Repurposing Video Models for 3D Super Resolution
Yuan Shen, Duygu Ceylan, Paul Guerrero +4
We present a simple, modular, and generic method that upsamples coarse 3D models by adding geometric and appearance details. While generative 3D models now exist, they do not yet m…
GRIP: Generating Interaction Poses Using Spatial Cues and Latent Consistency
Omid Taheri, Yi Zhou, Dimitrios Tzionas +4
Hands are dexterous and highly versatile manipulators that are central to how humans interact with objects and their environment. Consequently, modeling realistic hand-object inter…