9 citations · 15 across the 5 of their papers we have counts for
5 papers
GROUNDHOG: Grounding Large Language Models to Holistic Segmentation
Yichi Zhang, Ziqiao Ma, Xiaofeng Gao +3
Most multimodal large language models (MLLMs) learn language-to-object grounding through causal language modeling where grounded objects are captured by bounding boxes as sequences…
LEMMA: Learning Language-Conditioned Multi-Robot Manipulation
Ran Gong, Xiaofeng Gao, Qiaozi Gao +3
Complex manipulation tasks often require robots with complementary capabilities to collaborate. We introduce a benchmark for LanguagE-Conditioned Multi-robot MAnipulation (LEMMA) f…
Alexa, play with robot: Introducing the First Alexa Prize SimBot Challenge on Embodied AI
Hangjie Shi, Leslie Ball, Govind Thattai +39
The Alexa Prize program has empowered numerous university students to explore, experiment, and showcase their talents in building conversational agents through challenges like the…
Accelerator-Aware Training for Transducer-Based Speech Recognition
Suhaila M. Shakiah, Rupak Vignesh Swaminathan, Hieu Duy Nguyen +6
Machine learning model weights and activations are represented in full-precision during training. This leads to performance degradation in runtime when deployed on neural network a…
Alexa Arena: A User-Centric Interactive Platform for Embodied AI
Qiaozi Gao, Govind Thattai, Suhaila Shakiah +24
We introduce Alexa Arena, a user-centric simulation platform for Embodied AI (EAI) research. Alexa Arena provides a variety of multi-room layouts and interactable objects, for the…