3 papers
cs.CV2026
Can Language Models Learn to Listen?
Evonne Ng, Sanjay Subramanian, Dan Klein +3
We present a framework for generating appropriate facial responses from a listener in dyadic social interactions based on the speaker's words. Given an input transcription of the s…
cs.CV2026
Fast Image-based Neural Relighting with Translucency-Reflection Modeling
Shizhan Zhu, Shunsuke Saito, Aljaz Bozic +3
Image-based lighting (IBL) is a widely used technique that renders objects using a high dynamic range image or environment map. However, aggregating the irradiance at the object's…
cs.RO2025
Open X-Embodiment: Robotic Learning Datasets and RT-X Models
Embodiment Collaboration, Abby O'Neill, Abdul Rehman +291
Large, high-capacity models trained on diverse datasets have shown remarkable successes on efficiently tackling downstream applications. In domains from NLP to Computer Vision, thi…