11 papers
MicroZoom: Structure-Preserving Detail Synthesis at Extreme Scale
Huy Huynh, Jingwei Ma, Brian Curless +2
We introduce MicroZoom, a generative framework for gigapixel image synthesis at the microscopic scale. Given a standard photograph and a sparse set of consumer-grade microscope clo…
GarmentZoom: Generating Zoomable Images from Garment Listings
Renjie Zhao, Jingwei Ma, Huy Huynh Cao +3
Online product listings for garments often include an overview photo and a close-up to show garment details. However, each photo focuses on either field of view or garment detail,…
MusicInfuser: Making Video Diffusion Listen and Dance
Susung Hong, Ira Kemelmacher-Shlizerman, Brian Curless +1
We introduce MusicInfuser, an approach that aligns pre-trained text-to-video diffusion models to generate high-quality dance videos synchronized with specified music tracks. Rather…
COMIC: Agentic Sketch Comedy Generation
Susung Hong, Brian Curless, Ira Kemelmacher-Shlizerman +1
We propose a fully automated AI system that produces short comedic videos similar to sketch shows such as Saturday Night Live. Starting with character references, the system employ…
Prior-Enhanced Gaussian Splatting for Dynamic Scene Reconstruction from Casual Video
Meng-Li Shih, Ying-Huan Chen, Yu-Lun Liu +1
We introduce a fully automatic pipeline for dynamic scene reconstruction from casually captured monocular RGB videos. Rather than designing a new scene representation, we enhance t…
Generating Fit Check Videos with a Handheld Camera
Bowei Chen, Brian Curless, Ira Kemelmacher-Shlizerman +1
Self-captured full-body videos are popular, but most deployments require mounted cameras, carefully-framed shots, and repeated practice. We propose a more convenient solution that…