From the 1 of 18 linked papers with an AI index.
18 papers
X-Lens: Real-Time Metric Depth Estimation with Heterogeneous Cameras
Heng Zhou, Shuhong Liu, Yonghao He +6
X-Lens is a compact feed‑forward model that estimates metric depth in real time from a mix of calibrated fisheye and pinhole camera views using geometry‑aware calibration tokens an…
CapTalk: Text-Guided Stylization and Speech-Driven 3D Head Animation
Xuangeng Chu, Yuan Gan, Ziteng Cui +4
Audio-driven 3D facial animation aims to generate synchronized lip movements and vivid facial expressions from arbitrary audio clips. While existing methods can produce synchronize…
RAWild: Sensor-Agnostic RAW Object Detection via Physics-Guided Curve and Grid Modeling
Shuhong Liu, Gengjia Chang, Jun Liu +4
Camera sensor RAW data offers intrinsic advantages for object detection, including deeper bit depth, preserved physical information, and freedom from image signal processor (ISP) d…
FluxFlow: Conservative Flow-Matching for Astronomical Image Super-Resolution
Shuhong Liu, Xining Ge, Ziteng Cui +8
Ground-to-space astronomical super-resolution requires recovering space-quality images from ground-based observations that are simultaneously limited by pixel sampling resolution a…
NTIRE 2026 3D Restoration and Reconstruction in Real-world Adverse Conditions: RealX3D Challenge Results
Shuhong Liu, Chenyu Bao, Ziteng Cui +103
This paper presents a comprehensive review of the NTIRE 2026 3D Restoration and Reconstruction (3DRR) Challenge, detailing the proposed methods and results. The challenge seeks to…
Exploring Time Conditioning in Diffusion Generative Models from Disjoint Noisy Data Manifolds
Liuzhuozheng Li, Zhiyuan Zhan, Shuhong Liu +5
Practically, training diffusion models typically requires explicit time conditioning to guide the network through the denoising sampling process. Especially in deterministic method…