1 paper
Yongkun Du, Pinxuan Chen, Xuye Ying +1
The advent of Multimodal Large Language Models (MLLMs) has unlocked the potential for end-to-end document parsing and translation. However, prevailing benchmarks such as OmniDocBen…