Showing cs.MMShow all
3 papers · 1 filter
cs.MM2025
Hearing from Silence: Reasoning Audio Descriptions from Silent Videos via Vision-Language Model
Yong Ren, Chenxing Li, Le Xu +7
Humans can intuitively infer sounds from silent videos, but whether multimodal large language models can perform modal-mismatch reasoning without accessing target modalities remain…
cs.MM2025
Deconfounded Reasoning for Multimodal Fake News Detection via Causal Intervention
Moyang Liu, Kaiying Yan, Yukun Liu +4
The rapid growth of social media has led to the widespread dissemination of fake news across multiple content forms, including text, images, audio, and video. Traditional unimodal…
cs.MM2025
Exploring Modality Disruption in Multimodal Fake News Detection
Moyang Liu, Kaiying Yan, Yukun Liu +4
The rapid growth of social media has led to the widespread dissemination of fake news across multiple content forms, including text, images, audio, and video. Compared to unimodal…