2 citations · 2 across the 2 of their papers we have counts for
2 papers
cs.CV2025
Spatial Retrieval Augmented Autonomous Driving
Xiaosong Jia, Chenhe Zhang, Yule Jiang +8
Existing autonomous driving systems rely on onboard sensors (cameras, LiDAR, IMU, etc) for environmental perception. However, this paradigm is limited by the drive-time perception…
cs.CV2025★ 2 cited
MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing
Junbo Niu, Zheng Liu, Zhuangcheng Gu +58
We introduce MinerU2.5, a 1.2B-parameter document parsing vision-language model that achieves state-of-the-art recognition accuracy while maintaining exceptional computational effi…