7 papers
Thinking in Blender: Staged Executable Inverse Graphics with Vision-Language Models
Guangzhao He, Rundong Luo, Wei-Chiu Ma +1
Inverse graphics is a longstanding and highly underconstrained problem that seeks to reconstruct images as editable 3D scenes which can be rendered, relit, and manipulated. In this…
Composing People Together: Iterative Pose-Image Generation for Multi-Person Interaction Scenes
Wenxuan Peng, Bharath Hariharan, Hadar Averbuch-Elor
Despite recent progress, text-to-image models still struggle to generate semantically diverse and compositionally accurate multi-person interaction scenes, often collapsing to repe…
SceneAligner: 3D-Grounded Floorplan Localization in the Wild
Junhyeong Cho, Ruojin Cai, Hadar Averbuch-Elor
Many public buildings provide floorplans with a "you are here" indicator to help visitors orient themselves. Floorplan localization seeks to computationally replicate this capabili…
Raster2Seq: Polygon Sequence Generation for Floorplan Reconstruction
Hao Phung, Hadar Averbuch-Elor
Reconstructing a structured vector-graphics representation from a rasterized floorplan image is typically an important prerequisite for computational tasks involving floorplans suc…
Prox-E: Fine-Grained 3D Shape Editing via Primitive-Based Abstractions
Etai Sella, Hao Phung, Nitay Amiel +3
Text-based 2D image editing models have recently reached an impressive level of maturity, motivating a growing body of work that heavily depends on these models to drive 3D edits.…
MeshOn: Intersection-Free Mesh-to-Mesh Composition
Hyunwoo Kim, Itai Lang, Hadar Averbuch-Elor +2
We propose MeshOn, a method that finds physically and semantically realistic compositions of two input meshes. Given an accessory, a base mesh with a user-defined target region, an…