1 paper
Linhan Cao, Siyuan Li, Jun Lan +8
Large multimodal models (LMMs) have demonstrated strong OCR recognition capabilities, yet remain vulnerable to adversarial visual text that is readable to humans but challenging fo…