Showing cs.CVShow all
3 papers · 1 filter
cs.CV2025
MonSTeR: a Unified Model for Motion, Scene, Text Retrieval
Luca Collorone, Matteo Gioia, Massimiliano Pappa +5
Intention drives human movement in complex environments, but such movement can only happen if the surrounding context supports it. Despite the intuitive nature of this mechanism, e…
cs.CV2025
ANTHROPOS-V: benchmarking the novel task of Crowd Volume Estimation
Luca Collorone, Stefano D'Arrigo, Massimiliano Pappa +3
We introduce the novel task of Crowd Volume Estimation (CVE), defined as the process of estimating the collective body volume of crowds using only RGB images. Besides event managem…
cs.CV2024
MoDiPO: text-to-motion alignment via AI-feedback-driven Direct Preference Optimization
Massimiliano Pappa, Luca Collorone, Giovanni Ficarra +2
Diffusion Models have revolutionized the field of human motion generation by offering exceptional generation quality and fine-grained controllability through natural language condi…