From the 1 of 9 linked papers with an AI index.
8 papers
Scenix: Sparse-View 3D Scene Reconstruction via Executable Scene Programs
Kai Li, Lutao Jiang, Zhenyang Li +10
Synthesizing a structured and editable 3D indoor scene from a few uncalibrated RGB views requires more than generating high-quality individual assets: a system must infer the room…
ObliCity: A Benchmark and Baseline for Roof-to-Ground Projection Displacement Correction
Kai Li, Yupeng Deng, Ligao Deng +6
The paper presents ObliCity, a large-scale benchmark for extracting roof-to-footprint offset vectors in oblique urban remote sensing images, and introduces DragRoof, an ODE-based m…
DGTRSD & DGTRS-CLIP: A Dual-Granularity Remote Sensing Image-Text Dataset and Vision Language Foundation Model for Alignment
Weizhi Chen, Yupeng Deng, Jin Wei +7
Vision Language Foundation Models based on CLIP architecture for remote sensing primarily rely on short text captions, which often result in incomplete semantic representations. Al…
DragOSM: Extract Building Roofs and Footprints from Aerial Images by Aligning Historical Labels
Kai Li, Xingxing Weng, Yupeng Deng +4
Extracting polygonal roofs and footprints from remote sensing images is critical for large-scale urban analysis. Most existing methods rely on segmentation-based models that assume…
IRSAMap:Towards Large-Scale, High-Resolution Land Cover Map Vectorization
Yu Meng, Ligao Deng, Zhihao Xi +9
With the enhancement of remote sensing image resolution and the rapid advancement of deep learning, land cover mapping is transitioning from pixel-level segmentation to object-base…
PolyFootNet: Extracting Polygonal Building Footprints in Off-Nadir Remote Sensing Images
Kai Li, Yupeng Deng, Jingbo Chen +6
Extracting polygonal building footprints from off-nadir imagery is crucial for diverse applications. Current deep-learning-based extraction approaches predominantly rely on semanti…