1 citations · 1 across the 1 of their papers we have counts for
4 papers
TI-PREGO: Chain of Thought and In-Context Learning for Online Mistake Detection in PRocedural EGOcentric Videos
Leonardo Plini, Luca Scofano, Edoardo De Matteis +6
Identifying procedural errors online from egocentric videos is a critical yet challenging task across various domains, including manufacturing, healthcare, and skill-based training…
Compositional Entailment Learning for Hyperbolic Vision-Language Models
Avik Pal, Max van Spengler, Guido Maria D'Amely di Melendugno +3
Image-text representation learning forms a cornerstone in vision-language models, where pairs of images and textual descriptions are contrastively aligned in a shared embedding spa…
Hyp2Nav: Hyperbolic Planning and Curiosity for Crowd Navigation
Guido Maria D'Amely di Melendugno, Alessandro Flaborea, Pascal Mettes +1
Autonomous robots are increasingly becoming a strong fixture in social environments. Effective crowd navigation requires not only safe yet fast planning, but should also enable int…
PREGO: online mistake detection in PRocedural EGOcentric videos
Alessandro Flaborea, Guido Maria D'Amely di Melendugno, Leonardo Plini +5
Promptly identifying procedural errors from egocentric videos in an online setting is highly challenging and valuable for detecting mistakes as soon as they happen. This capability…