activity
20232026
most citedRoboCLIP: One Demonstration is Enough to Learn Robot Policies

7 citations · 10 across the 3 of their papers we have counts for

collaborators

11 papers

cs.CV2026

Three-Step Nav: A Hierarchical Global-Local Planner for Zero-Shot Vision-and-Language Navigation

Wanrong Zheng, Yunhao Ge, Laurent Itti

Breakthrough progress in vision-based navigation through unknown environments has been achieved by using multimodal large language models (MLLMs). These models can plan a sequence…

cs.RO2025

Towards Embodiment Scaling Laws in Robot Locomotion

Bo Ai, Liu Dai, Nico Bohlinger +7

Cross-embodiment generalization underpins the vision of building generalist embodied agents for any robot, yet its enabling factors remain poorly understood. We investigate embodim…

cs.NE2025

Perforated Backpropagation: A Neuroscience Inspired Extension to Artificial Neural Networks

Rorry Brenner, Laurent Itti

The neurons of artificial neural networks were originally invented when much less was known about biological neurons than is known today. Our work explores a modification to the co…

cs.CV2024

BEHAVIOR Vision Suite: Customizable Dataset Generation via Simulation

Yunhao Ge, Yihe Tang, Jiashu Xu +20

The systematic evaluation and understanding of computer vision models under varying conditions require large amounts of data with comprehensive and customized labels, which real-wo…

cs.LG2023

Evaluating Pretrained models for Deployable Lifelong Learning

Kiran Lekkala, Eshan Bhargava, Yunhao Ge +1

We create a novel benchmark for evaluating a Deployable Lifelong Learning system for Visual Reinforcement Learning (RL) that is pretrained on a curated dataset, and propose a novel…

cs.CV2023

3D Copy-Paste: Physically Plausible Object Insertion for Monocular 3D Detection

Yunhao Ge, Hong-Xing Yu, Cheng Zhao +5

A major challenge in monocular 3D object detection is the limited diversity and quantity of objects in real datasets. While augmenting real scenes with virtual objects holds promis…