From the 2 of 71 linked papers with an AI index.
49 papers · 1 filter
Map-Det3D: Metric Feed-Forward 3D Reconstruction Prior for Multi-view 3D Object Detection from Streaming Inputs
Yung-Hsu Yang, Luigi Piccinelli, Samuel Rota Bulò +7
Metric 3D object detection is a core capability for embodied agents, yet most reliable systems lean on depth sensors, trading away cost, power, and integration simplicity. This mot…
EgoTrack3D: A Modular Framework for Egocentric 3D Object Tracking
Jan Kulik, Bjarni Dagur Thor Karason, Yung-Hsu Yang +3
Understanding 3D scenes from egocentric video is fundamental for robotics and autonomous navigation, yet rapid viewpoint changes and partial occlusions make building structured rep…
DVPSFormer: Efficient Online Depth-aware Video Panoptic Segmentation for Autonomous Driving
Yung-Hsu Yang, Luigi Piccinelli, Siyuan Li +8
DVPSFormer is an online architecture that jointly estimates metric depth, semantic segmentation, and instance trajectories for autonomous driving by using explicit scene discretiza…
Head Avatars with Dynamic Explicit Hair
Vanessa Sklyarova, Haonan Chen, Berna Kabadayi +8
We present DynHair, a novel method for tracking and modeling dynamic hair for human head avatars. From video input, we reconstruct a dynamic head avatar with an explicit strand-bas…
LangLoc: "Tell Me What You See"
Shaurya Kishore Panwar, Roham Zendehdel Nobari, Shirley Feng Yi Lau +4
We tackle fine-grained indoor localization from natural language: given a free-form description of one's surroundings, estimate the observer's 2D position and heading within a know…
SuperFlex: Deformable Superquadrics for Point Cloud Decomposition
Gabriel Tavernini, Elisabetta Fedele, Tiago Novello +3
Superquadrics have proven to provide a compact, geometrically meaningful representation for 3D objects. However, existing methods suffer from limited reconstruction accuracy, are r…