2 papers
cs.CV2026
PVLM: Parsing-Aware Vision Language Model with Dynamic Contrastive Learning for Zero-Shot Deepfake Attribution
Yaning Zhang, Jiahe Zhang, Chunjie Ma +3
The challenge of tracing the source attribution of forged faces has gained significant attention due to the rapid advancement of generative models. However, existing deepfake attri…
cs.CV2026
SpatialV2A: Visual-Guided High-fidelity Spatial Audio Generation
Yanan Wang, Linjie Ren, Zihao Li +2
While video-to-audio generation has achieved remarkable progress in semantic and temporal alignment, most existing studies focus solely on these aspects, paying limited attention t…