1 paper · 1 filter
Harikrishnan P M, Goutham Vignesh, Ganesh Parab +4
Efficient multimodal document question answering with explicit visual grounding, locating the precise document region that supports each answer remains an open challenge. Current a…