3 papers
cs.CV2026
ReViCo: Unveiling the Limitations of VLMs in Visual Text Understanding via Error Correction
Bojun Zhang, Junhong Liang, Feifei Zhai +2
Vision Language Models (VLMs) have shown great success in general visual tasks, yet they still struggle to deeply understand text within images. In this paper, we introduce ReViCo…
cs.CL2026
Listen, Correct, and Feed Back: Spoken Pedagogical Feedback Generation
Junhong Liang, Yifan Lu, Ekaterina Kochmar +1
Grammatical error correction (GEC) and explanation (GEE) have made rapid progress, but real teaching scenarios also require \emph{learner-friendly pedagogical feedback} that is act…
cs.CL2025
Vision Language Models Are Not (Yet) Spelling Correctors
Junhong Liang, Bojun Zhang
Spelling correction from visual input poses unique challenges for vision language models (VLMs), as it requires not only detecting but also correcting textual errors directly withi…