2 papers
cs.LG2025
Sample what you cant compress
Vighnesh Birodkar, Gabriel Barcik, James Lyon +3
For learned image representations, basic autoencoders often produce blurry results. Reconstruction quality can be improved by incorporating additional penalties such as adversarial…
cs.CV2024
Learning Complex Non-Rigid Image Edits from Multimodal Conditioning
Nikolai Warner, Jack Kolb, Meera Hahn +3
In this paper we focus on inserting a given human (specifically, a single image of a person) into a novel scene. Our method, which builds on top of Stable Diffusion, yields natural…