3 papers
cs.CV2023
FineControlNet: Fine-level Text Control for Image Generation with Spatially Aligned Text Control Injection
Hongsuk Choi, Isaac Kasahara, Selim Engin +3
Recently introduced ControlNet has the ability to steer the text-driven image generation process with geometric input such as human 2D pose, or edge features. While ControlNet prov…
cs.CV2023
VioLA: Aligning Videos to 2D LiDAR Scans
Jun-Jee Chao, Selim Engin, Nikhil Chavan-Dafle +2
We study the problem of aligning a video that captures a local portion of an environment to the 2D LiDAR scan of the entire environment. We introduce a method (VioLA) that starts w…
cs.CV2023
HandNeRF: Learning to Reconstruct Hand-Object Interaction Scene from a Single RGB Image
Hongsuk Choi, Nikhil Chavan-Dafle, Jiacheng Yuan +2
This paper presents a method to learn hand-object interaction prior for reconstructing a 3D hand-object scene from a single RGB image. The inference as well as training-data genera…