3 papers
cs.CV2026
GeoFF3D: Coordinate-Anchored Feed-Forward Reconstruction for Large-Scale UAV Mapping
Xiang Yang, Yongli Wang, Yunsheng Zhang +3
Existing feed-forward 3D reconstruction methods typically process a bounded number of images and recover cameras and geometry in local or internally normalized frames. Extending th…
cs.CV2026
UAVFF3D: A Geometry-Aware Benchmark for Feed-Forward UAV 3D Reconstruction
Xiang Yang, Yongli Wang, HaiFeng Li +1
Feed-forward 3D reconstruction has advanced rapidly, but current models remain unreliable in UAV photogrammetric acquisition. We argue that this failure is caused not only by appea…
cs.CV2024
Homogeneous Tokenizer Matters: Homogeneous Visual Tokenizer for Remote Sensing Image Understanding
Run Shao, Zhaoyang Zhang, Chao Tao +3
The tokenizer, as one of the fundamental components of large models, has long been overlooked or even misunderstood in visual tasks. One key factor of the great comprehension power…