activity
20142025
most citedARMANI: Part-level Garment-Text Alignment for Unified Cross-Modal Fashion Design

27 citations · 205 across the 69 of their papers we have counts for

collaborators
Showing cs.ROShow all

5 papers · 1 filter

cs.RO2025

Embodied Arena: A Comprehensive, Unified, and Evolving Evaluation Platform for Embodied AI

Fei Ni, Min Zhang, Pengyi Li +34

Embodied AI development significantly lags behind large foundation models due to three critical challenges: (1) lack of systematic understanding of core capabilities needed for Emb…

cs.RO20241 cited

InfiniteWorld: A Unified Scalable Simulation Framework for General Visual-Language Robot Interaction

Pengzhen Ren, Min Li, Zhen Luo +15

Realizing scaling laws in embodied AI has become a focus. However, previous work has been scattered across diverse simulation platforms, with assets and models lacking unified inte…

cs.RO20241 cited

InstruGen: Automatic Instruction Generation for Vision-and-Language Navigation Via Large Multimodal Models

Yu Yan, Rongtao Xu, Jiazhao Zhang +3

Recent research on Vision-and-Language Navigation (VLN) indicates that agents suffer from poor generalization in unseen environments due to the lack of realistic training environme…

cs.RO20242 cited

All Robots in One: A New Standard and Unified Dataset for Versatile, General-Purpose Embodied Agents

Zhiqiang Wang, Hao Zheng, Yunshuang Nie +11

Embodied AI is transforming how AI systems interact with the physical world, yet existing datasets are inadequate for developing versatile, general-purpose agents. These limitation…

cs.RO2024

Affordances-Oriented Planning using Foundation Models for Continuous Vision-Language Navigation

Jiaqi Chen, Bingqian Lin, Xinmin Liu +3

LLM-based agents have demonstrated impressive zero-shot performance in vision-language navigation (VLN) task. However, existing LLM-based methods often focus only on solving high-l…