10 papers
MicroZoom: Structure-Preserving Detail Synthesis at Extreme Scale
Huy Huynh, Jingwei Ma, Brian Curless +2
We introduce MicroZoom, a generative framework for gigapixel image synthesis at the microscopic scale. Given a standard photograph and a sparse set of consumer-grade microscope clo…
GarmentZoom: Generating Zoomable Images from Garment Listings
Renjie Zhao, Jingwei Ma, Huy Huynh Cao +3
Online product listings for garments often include an overview photo and a close-up to show garment details. However, each photo focuses on either field of view or garment detail,…
MusicInfuser: Making Video Diffusion Listen and Dance
Susung Hong, Ira Kemelmacher-Shlizerman, Brian Curless +1
We introduce MusicInfuser, an approach that aligns pre-trained text-to-video diffusion models to generate high-quality dance videos synchronized with specified music tracks. Rather…
COMIC: Agentic Sketch Comedy Generation
Susung Hong, Brian Curless, Ira Kemelmacher-Shlizerman +1
We propose a fully automated AI system that produces short comedic videos similar to sketch shows such as Saturday Night Live. Starting with character references, the system employ…
Generating Fit Check Videos with a Handheld Camera
Bowei Chen, Brian Curless, Ira Kemelmacher-Shlizerman +1
Self-captured full-body videos are popular, but most deployments require mounted cameras, carefully-framed shots, and repeated practice. We propose a more convenient solution that…
Linearly Constrained Diffusion Implicit Models
Vivek Jayaram, Ira Kemelmacher-Shlizerman, Steven M. Seitz +1
We introduce Linearly Constrained Diffusion Implicit Models (CDIM), a fast and accurate approach to solving noisy linear inverse problems using diffusion models. Traditional diffus…