works on

From the 1 of 5 linked papers with an AI index.

collaborators

5 papers

cs.CV2026

ViCo3D: Empowering LiDAR-based Collaborative 3D Object Detection with Vision Foundation Models

Haojie Ren, Songrui Luo, Lingfeng Wang +6

The paper introduces ViCo3D, a framework that leverages vision foundation models to enrich LiDAR bird's-eye-view features for collaborative 3D object detection in V2X scenarios, ac…

cs.RO2026

CAC-VLA: Context-Gated Action Conditioning for Vision-Language-Action Models

Yifu Xiong, Wenhao Yu, Jiaxuan Lin +5

Vision-Language-Action (VLA) models have become a promising paradigm for generalist robot manipulation, where visual-language representations are used to condition continuous actio…

cs.CV2026

CORP: A Multi-Modal Dataset for Campus-Oriented Roadside Perception Tasks

Beibei Wang, Zijian Yu, Lu Zhang +8

Numerous roadside perception datasets have been introduced to propel advancements in autonomous driving and intelligent transportation systems research and development. However, it…

cs.AI2025

\(X\)-evolve: Solution space evolution powered by large language models

Yi Zhai, Zhiqiang Wei, Ruohan Li +7

While combining large language models (LLMs) with evolutionary algorithms (EAs) shows promise for solving complex optimization problems, current approaches typically evolve individ…

cs.RO2025

MT-PCR: Leveraging Modality Transformation for Large-Scale Point Cloud Registration with Limited Overlap

Yilong Wu, Yifan Duan, Yuxi Chen +5

Large-scale scene point cloud registration with limited overlap is a challenging task due to computational load and constrained data acquisition. To tackle these issues, we propose…