565 citations · 1.4k across the 55 of their papers we have counts for
Showing 2026 · cs.CVShow all
2 papers · 2 filters
cs.CV2026
Hyper3-CLIP: Hierarchy-Conditioned Hyperbolic Vision-Language Training
Matin Mahmood, Antonio Rueda-Toicen, Mohamed ElBassat +3
CLIP-like vision-language models (VLMs) trained with contrastive objectives learn strong global image-text representations, but their Euclidean embeddings and global pooling fail t…
cs.CV2026
Are We Recognizing the Jaguar or Its Background? A Diagnostic Framework for Jaguar Re-Identification
Antonio Rueda-Toicen, Abigail Allen Martin, Daniil Morozov +5
Jaguar re-identification (re-ID) from citizen-science imagery can look strong on standard retrieval metrics while still relying on the wrong evidence, such as background context or…