7 citations · 10 across the 3 of their papers we have counts for
11 papers
Three-Step Nav: A Hierarchical Global-Local Planner for Zero-Shot Vision-and-Language Navigation
Wanrong Zheng, Yunhao Ge, Laurent Itti
Breakthrough progress in vision-based navigation through unknown environments has been achieved by using multimodal large language models (MLLMs). These models can plan a sequence…
Towards Embodiment Scaling Laws in Robot Locomotion
Bo Ai, Liu Dai, Nico Bohlinger +7
Cross-embodiment generalization underpins the vision of building generalist embodied agents for any robot, yet its enabling factors remain poorly understood. We investigate embodim…
Perforated Backpropagation: A Neuroscience Inspired Extension to Artificial Neural Networks
Rorry Brenner, Laurent Itti
The neurons of artificial neural networks were originally invented when much less was known about biological neurons than is known today. Our work explores a modification to the co…
BEHAVIOR Vision Suite: Customizable Dataset Generation via Simulation
Yunhao Ge, Yihe Tang, Jiashu Xu +20
The systematic evaluation and understanding of computer vision models under varying conditions require large amounts of data with comprehensive and customized labels, which real-wo…
Evaluating Pretrained models for Deployable Lifelong Learning
Kiran Lekkala, Eshan Bhargava, Yunhao Ge +1
We create a novel benchmark for evaluating a Deployable Lifelong Learning system for Visual Reinforcement Learning (RL) that is pretrained on a curated dataset, and propose a novel…
3D Copy-Paste: Physically Plausible Object Insertion for Monocular 3D Detection
Yunhao Ge, Hong-Xing Yu, Cheng Zhao +5
A major challenge in monocular 3D object detection is the limited diversity and quantity of objects in real datasets. While augmenting real scenes with virtual objects holds promis…