54 citations · 92 across the 20 of their papers we have counts for
23 papers · 1 filter
ForeSplat: Optimization-Aware Foresight for Feed-Forward 3D Gaussian Splatting
Yuke Li, Weihang Liu, Cheng Zhang +8
Feed-forward 3D Gaussian Splatting models offer fast single-pass reconstruction,but scaling them to match per-scene optimization quality is fundamentally hindered by the scarcity o…
CADSpotting: Robust Panoptic Symbol Spotting on Large-Scale CAD Drawings
Fuyi Yang, Jiazuo Mu, Yanshun Zhang +7
We introduce CADSpotting, an effective method for panoptic symbol spotting in large-scale architectural CAD drawings. Existing approaches often struggle with symbol diversity, scal…
AerialGo: Walking-through City View Generation from Aerial Perspectives
Fuqiang Zhao, Yijing Guo, Siyuan Yang +6
High-quality 3D urban reconstruction is essential for applications in urban planning, navigation, and AR/VR. However, capturing detailed ground-level data across cities is both lab…
V^3: Viewing Volumetric Videos on Mobiles via Streamable 2D Dynamic Gaussians
Penghao Wang, Zhirui Zhang, Liao Wang +5
Experiencing high-fidelity volumetric video as seamlessly as 2D videos is a long-held dream. However, current dynamic 3DGS methods, despite their high rendering quality, face chall…
SCOPE: Sign Language Contextual Processing with Embedding from LLMs
Yuqi Liu, Wenqian Zhang, Sihan Ren +3
Sign languages, used by around 70 million Deaf individuals globally, are visual languages that convey visual and contextual information. Current methods in vision-based sign langua…
StackFLOW: Monocular Human-Object Reconstruction by Stacked Normalizing Flow with Offset
Chaofan Huo, Ye Shi, Yuexin Ma +3
Modeling and capturing the 3D spatial arrangement of the human and the object is the key to perceiving 3D human-object interaction from monocular images. In this work, we propose t…