2 papers
cs.CV2026
SalsaAgent: A multimodal embodied language model for interactive dance generation
Payam Jome Yazdian, Zoe Stanley, Angelica Lim
Embodied interaction with humanoids depends on bidirectional nonverbal reactivity, coordination, and synchrony to convey cues and move with a partner. For socially interactive embo…
cs.CV2026
Chehre: An Emoji-Prompted Dataset to Explore Perceptual Flexibility in Video Language Models
Bita Azari, Zoe Stanley, Avneet Batra +4
Do people perceive the same facial expression in the same way? Should we expect vision models to be flexible in how they perceive facial expressions? Facial expressions are nonverb…