most citedGhost in the Minecraft: Generally Capable Agents for Open-World Environments via Large Language Models with Text-based Knowledge and Memory

24 citations · 35 across the 6 of their papers we have counts for

collaborators

6 papers

cs.CV2023

Tracking Objects with 3D Representation from Videos

Jiawei He, Lue Fan, Yuqi Wang +4

Data association is a knotty problem for 2D Multiple Object Tracking due to the object occlusion. However, in 3D space, data association is not so hard. Only with a 3D Kalman Filte…

cs.AI202324 cited

Ghost in the Minecraft: Generally Capable Agents for Open-World Environments via Large Language Models with Text-based Knowledge and Memory

Xizhou Zhu, Yuntao Chen, Hao Tian +10

The captivating realm of Minecraft has attracted substantial research interest in recent years, serving as a rich platform for developing intelligent agents capable of functioning…

cs.CV20233 cited

Real-Aug: Realistic Scene Synthesis for LiDAR Augmentation in 3D Object Detection

Jinglin Zhan, Tiejun Liu, Rengang Li +3

Data and model are the undoubtable two supporting pillars for LiDAR object detection. However, data-centric works have fallen far behind compared with the ever-growing list of fanc…

cs.CV20236 cited

Once Detected, Never Lost: Surpassing Human Performance in Offline LiDAR based 3D Object Detection

Lue Fan, Yuxue Yang, Yiming Mao +4

This paper aims for high-performance offline LiDAR-based 3D object detection. We first observe that experienced human annotators annotate objects from a track-centric perspective.…

cs.CV20231 cited

3D Video Object Detection with Learnable Object-Centric Global Optimization

Jiawei He, Yuntao Chen, Naiyan Wang +1

We explore long-term temporal visual correspondence-based optimization for 3D video object detection in this work. Visual correspondence refers to one-to-one mappings for pixels ac…

cs.CV20221 cited

Densely Constrained Depth Estimator for Monocular 3D Object Detection

Yingyan Li, Yuntao Chen, Jiawei He +1

Estimating accurate 3D locations of objects from monocular images is a challenging problem because of lacking depth. Previous work shows that utilizing the object's keypoint projec…