1 paper · 1 filter
Yuchen Deng, Chang Sun, Hai-Tao Zheng +2
Omnimodal large language models (Omni-LLMs) integrate audio, video, and text, yet remain vulnerable to cross-modal hallucinations, where one modality improperly influences predicti…