3 papers
cs.LG2026
LEVANTE-bench: Multi-Scale Comparison of VLMs to Children Using Cognitive Tasks (or, "Is Your VLM Smarter Than a 5th Grader?")
Alvin Wei Ming Tan, David Cardinal, Tania Lorido-Botran +3
Given the inherently multimodal nature of human experience, vision-language models (VLMs) hold substantial promise for modeling human cognition as it grows and develops with experi…
cs.CV2026
Anny-Fit: All-Age Human Mesh Recovery
Laura Bravo-Sánchez, Matthieu Armando, Romain Brégier +3
Recovering 3D human pose and shape from a single image remains a cornerstone of human-centric vision, yet most methods assume adult subjects and optimize each person independently.…
cs.CV2025
Human Mesh Modeling for Anny Body
Romain Brégier, Guénolé Fiche, Laura Bravo-Sánchez +5
Parametric body models provide the structural basis for many human-centric tasks, yet existing models often rely on costly 3D scans and learned shape spaces that are proprietary an…