3d representation learning 1euclidean transformations 1latent space navigation 1self-supervised vision 1visual odometry 1
From the 1 of 3 linked papers with an AI index.
3 papers
cs.CV2026
Primitive-Driven Compositional Forensic Visual Prompting for Open-World Face Anti-Spoofing
Fangling Jiang, Qi Li, Bing Liu +4
Open-world face anti-spoofing must address both covariate and semantic shifts: source and target domains differ in imaging conditions, while target domains contain diverse attack t…
cs.CV2026
SeeSE3: Emergence of 3D Space in Vision Features
Caroline Chen, Sayna Ebrahimi, Fedor Kitashov +4
The paper examines whether vision foundation models implicitly encode the geometry of 3D Euclidean space, introducing probes such as a mutual neighborhood metric and a Poincaré Ada…
cs.CV2024
Structured Video-Language Modeling with Temporal Grouping and Spatial Grounding
Yuanhao Xiong, Long Zhao, Boqing Gong +5
Existing video-language pre-training methods primarily focus on instance-level alignment between video clips and captions via global contrastive learning but neglect rich fine-grai…