From the 1 of 7 linked papers with an AI index.
7 papers
Pix2Act: Image-Space Manipulation Policies with Equivariant Augmentation
Haojie Huang, Linfeng Zhao, Haotian Liu +9
Pix2Act is an imitation‑learning approach that predicts continuous 2D keypoint trajectories in camera images and recovers 3D end‑effector poses via triangulation, using equivariant…
Pantheon360: Taming Digital Twin Generation via 3D-Aware 360° Video Diffusion
Ting-Hsuan Chen, Ying-Huan Chen, Tao Tu +10
Generating complete digital twins from videos requires precise camera control, global scene coverage, and strict spatial-temporal consistency constraints that remain challenging fo…
A Strong View-Free Baseline Approach for Single-View Image Guided Point Cloud Completion
Fangzhou Lin, Zilin Dai, Rigved Sanku +4
The single-view image guided point cloud completion (SVIPC) task aims to reconstruct a complete point cloud from a partial input with the help of a single-view image. While previou…
Mobile Augmented Reality Framework with Fusional Localization and Pose Estimation
Songlin Hou, Fangzhou Lin, Yunmei Huang +2
As a novel way of presenting information, augmented reality (AR) enables people to interact with the physical world in a direct and intuitive way. While there are some mobile AR pr…
Hyperbolic Chamfer Distance for Point Cloud Completion and Beyond
Fangzhou Lin, Songlin Hou, Haotian Liu +4
Chamfer Distance (CD) is widely used as a metric to quantify difference between two point clouds. In point cloud completion, Chamfer Distance (CD) is typically used as a loss funct…
SPADE: Spectroscopic Photoacoustic Denoising using an Analytical and Data-free Enhancement Framework
Fangzhou Lin, Shang Gao, Yichuan Tang +6
Spectroscopic photoacoustic (sPA) imaging uses multiple wavelengths to differentiate chromophores based on their unique optical absorption spectra. This technique has been widely a…