1 paper · 1 filter
Yen-Ting Piao, Shu-Yun Chen, Chin-Hui Chu +4
Omni-modal large language models (OLLMs) jointly process vision, audio, and text, yet their modality bias under cross-modal conflict remains underexplored. Existing benchmarks conf…