22 citations · 76 across the 11 of their papers we have counts for
11 papers
MinerU: An Open-Source Solution for Precise Document Content Extraction
Bin Wang, Chao Xu, Xiaomeng Zhao +15
Document content analysis has been a crucial research area in computer vision. Despite significant advancements in methods such as OCR, layout detection, and formula recognition, e…
OASim: an Open and Adaptive Simulator based on Neural Rendering for Autonomous Driving
Guohang Yan, Jiahao Pi, Jianfei Guo +14
With deep learning and computer vision technology development, autonomous driving provides new solutions to improve traffic safety and efficiency. The importance of building high-q…
LimSim++: A Closed-Loop Platform for Deploying Multimodal LLMs in Autonomous Driving
Daocheng Fu, Wenjie Lei, Licheng Wen +5
The emergence of Multimodal Large Language Models ((M)LLMs) has ushered in new avenues in artificial intelligence, particularly for autonomous driving by offering enhanced understa…
Towards Knowledge-driven Autonomous Driving
Xin Li, Yeqi Bai, Pinlong Cai +14
This paper explores the emerging knowledge-driven autonomous driving technologies. Our investigation highlights the limitations of current autonomous driving systems, in particular…
Drive Like a Human: Rethinking Autonomous Driving with Large Language Models
Daocheng Fu, Xin Li, Licheng Wen +4
In this paper, we explore the potential of using a large language model (LLM) to understand the driving environment in a human-like manner and analyze its ability to reason, interp…
StreetSurf: Extending Multi-view Implicit Surface Reconstruction to Street Views
Jianfei Guo, Nianchen Deng, Xinyang Li +6
We present a novel multi-view implicit surface reconstruction technique, termed StreetSurf, that is readily applicable to street view images in widely-used autonomous driving datas…