From the 1 of 10 linked papers with an AI index.
10 papers
SphereVideo: Prototype-anchored Hyperspherical Boundary for Continual AI-generated Video Detection
Fei Li, Yue Yu, Yuran Wang +3
AI-generated video (AIGV) detection aims to distinguish real videos from AI-generated ones. In practice, detectors trained on existing data often fail to generalize to newly emergi…
DECODE: Tackling Representation and Decision Degradation in Continual AI-Generated Image Detection
Zihao Cai, Xinghan Li, Ruiyan Yang +3
The paper introduces DECODE, a framework that addresses both representation and decision-level forgetting in continual learning for AI-generated image detection, using subspace div…
VIGIL: Part-Grounded Structured Reasoning for Generalizable Deepfake Detection
Xinghan Li, Junhao Xu, Jingjing Chen
Multimodal large language models (MLLMs) offer a promising path toward interpretable deepfake detection by generating textual explanations. However, the reasoning process of curren…
From Spatial to Actions: Grounding Vision-Language-Action Model in Spatial Foundation Priors
Zhengshen Zhang, Hao Li, Yalun Dai +10
Existing vision-language-action (VLA) models act in 3D real-world but are typically built on 2D encoders, leaving a spatial reasoning gap that limits generalization and adaptabilit…
Revealing the Implicit Noise-based Imprint of Generative Models
Xinghan Li, Yue Yu, Xue Song +2
With the rapid advancement of vision generation models, the potential security risks stemming from synthetic visual content have garnered increasing attention, posing significant c…
Adam Reduces a Unique Form of Sharpness: Theoretical Insights Near the Minimizer Manifold
Xinghan Li, Haodong Wen, Kaifeng Lyu
Despite the popularity of the Adam optimizer in practice, most theoretical analyses study Stochastic Gradient Descent (SGD) as a proxy for Adam, and little is known about how the s…