most citedIRSAMap:Towards Large-Scale, High-Resolution Land Cover Map Vectorization

3 citations · 3 across the 3 of their papers we have counts for

collaborators

5 papers

cs.CV2025

DragOSM: Extract Building Roofs and Footprints from Aerial Images by Aligning Historical Labels

Kai Li, Xingxing Weng, Yupeng Deng +4

Extracting polygonal roofs and footprints from remote sensing images is critical for large-scale urban analysis. Most existing methods rely on segmentation-based models that assume…

cs.CV20253 cited

IRSAMap:Towards Large-Scale, High-Resolution Land Cover Map Vectorization

Yu Meng, Ligao Deng, Zhihao Xi +9

With the enhancement of remote sensing image resolution and the rapid advancement of deep learning, land cover mapping is transitioning from pixel-level segmentation to object-base…

cs.CV2025

GLD-Road:A global-local decoding road network extraction model for remote sensing images

Ligao Deng, Yupeng Deng, Yu Meng +4

Road networks are crucial for mapping, autonomous driving, and disaster response. While manual annotation is costly, deep learning offers efficient extraction. Current methods incl…

cs.CV2025

DGTRSD & DGTRS-CLIP: A Dual-Granularity Remote Sensing Image-Text Dataset and Vision Language Foundation Model for Alignment

Weizhi Chen, Yupeng Deng, Jin Wei +7

Vision Language Foundation Models based on CLIP architecture for remote sensing primarily rely on short text captions, which often result in incomplete semantic representations. Al…

cs.CV2025

SayAnything: Audio-Driven Lip Synchronization with Conditional Video Diffusion

Junxian Ma, Shiwen Wang, Jian Yang +6

Recent advances in diffusion models have led to significant progress in audio-driven lip synchronization. However, existing methods typically rely on constrained audio-visual align…