1 paper
Xudong Liu, Bicheng Wan, Yulin Jin
End-to-end OCR systems based on vision-language models have achieved strong performance in complex document OCR, but their efficiency is limited by the large number of visual token…