6 citations · 8 across the 3 of their papers we have counts for
Showing cs.MMShow all
2 papers · 1 filter
cs.MM2026
AVID: A Benchmark for Omni-Modal Audio-Visual Inconsistency Understanding via Agent-Driven Construction
Zixuan Chen, Depeng Wang, Hao Lin +6
We present AVID, the first large-scale benchmark for audio-visual inconsistency understanding in videos. While omni-modal large language models excel at temporally aligned tasks su…
cs.MM2020★ 6 cited
Coverless Video Steganography based on Maximum DC Coefficients
Laijin Meng, Xinghao Jiang, Zhenzhen Zhang +2
Coverless steganography has been a great interest in recent years, since it is a technology that can absolutely resist the detection of steganalysis by not modifying the carriers.…