7 papers
HumanOrbit: 3D Human Reconstruction as 360° Orbit Generation
Keito Suzuki, Kunyao Chen, Lei Wang +5
We present a method for generating a full 360° orbit video around a person from a single input image. Existing methods typically adapt image-based diffusion models for multi-view…
SplatSDF: Boosting SDF-NeRF via Architecture-Level Fusion with Gaussian Splats
Runfa Blark Li, Keito Suzuki, Bang Du +3
Signed distance-radiance field (SDF-NeRF) is a promising environment representation that offers both photo-realistic rendering and geometric reasoning such as proximity queries for…
AirHunt: Bridging VLM Semantics and Continuous Planning for Efficient Aerial Object Navigation
Xuecheng Chen, Zongzhuo Liu, Jianfa Ma +4
Recent advances in large Vision-Language Models (VLMs) have provided rich semantic understanding that empowers drones to search for open-set objects via natural language instructio…
OpenHuman4D: Open-Vocabulary 4D Human Parsing
Keito Suzuki, Bang Du, Runfa Blark Li +5
Understanding dynamic 3D human representation has become increasingly critical in virtual and extended reality applications. However, existing human part segmentation methods are c…
DynaGSLAM: Real-Time Gaussian-Splatting SLAM for Online Rendering, Tracking, Motion Predictions of Moving Objects in Dynamic Scenes
Runfa Blark Li, Mahdi Shaghaghi, Keito Suzuki +8
Simultaneous Localization and Mapping (SLAM) is one of the most important environment-perception and navigation algorithms for computer vision, robotics, and autonomous cars/drones…
Open-Vocabulary Semantic Part Segmentation of 3D Human
Keito Suzuki, Bang Du, Girish Krishnan +3
3D part segmentation is still an open problem in the field of 3D vision and AR/VR. Due to limited 3D labeled data, traditional supervised segmentation methods fall short in general…