Showing cs.CVShow all
2 papers · 1 filter
cs.CV2025
RynnEC: Bringing MLLMs into Embodied World
Ronghao Dang, Yuqian Yuan, Yunxuan Mao +6
We introduce RynnEC, a video multimodal large language model designed for embodied cognition. Built upon a general-purpose vision-language foundation model, RynnEC incorporates a r…
cs.CV2024
Are Dense Labels Always Necessary for 3D Object Detection from Point Cloud?
Chenqiang Gao, Chuandong Liu, Jun Shu +5
Current state-of-the-art (SOTA) 3D object detection methods often require a large amount of 3D bounding box annotations for training. However, collecting such large-scale densely-s…