2 papers
cs.CV2026
OP-CAD: On-Policy Clean-Audio Distillation for Robust Audio-Visual Reasoning
Xingming Shui, Dapeng Chen, Bowei Liu +5
Omni-modal large language models deployed in real-world environments encounter external noise that can interfere with their perception and understanding of multimodal inputs. We st…
cs.CV2026
VidForensics-M1: Meta-Detection Reinforcement Learning with Verifiable Temporal Grounding for AI-Generated Video Forensics
Bowei Liu, Zheng Lu, Yuhan Bian +8
Recent advances in video generation models have significantly improved the realism of synthetic videos, blurring the boundary between generated and authentic content and raising co…