Showing cs.CVShow all
2 papers · 1 filter
cs.CV2025
Enhancing Spatial Reasoning through Visual and Textual Thinking
Xun Liang, Xin Guo, Zhongming Jin +5
The spatial reasoning task aims to reason about the spatial relationships in 2D and 3D space, which is a fundamental capability for Visual Question Answering (VQA) and robotics. Al…
cs.CV2021
Discriminative-Generative Dual Memory Video Anomaly Detection
Xin Guo, Zhongming Jin, Chong Chen +5
Recently, people tried to use a few anomalies for video anomaly detection (VAD) instead of only normal data during the training process. A side effect of data imbalance occurs when…