6 citations · 10 across the 3 of their papers we have counts for
3 papers
cs.CL2023★ 1 cited
ChatSpot: Bootstrapping Multimodal LLMs via Precise Referring Instruction Tuning
Liang Zhao, En Yu, Zheng Ge +8
Human-AI interactivity is a critical aspect that reflects the usability of multimodal large language models (MLLMs). However, existing end-to-end MLLMs only allow users to interact…
cs.CV2022★ 3 cited
PersDet: Monocular 3D Detection in Perspective Bird's-Eye-View
Hongyu Zhou, Zheng Ge, Weixin Mao +1
Currently, detecting 3D objects in Bird's-Eye-View (BEV) is superior to other 3D detectors for autonomous driving and robotics. However, transforming image features into BEV necess…
cs.CV2022★ 6 cited
Dense Teacher: Dense Pseudo-Labels for Semi-supervised Object Detection
Hongyu Zhou, Zheng Ge, Songtao Liu +4
To date, the most powerful semi-supervised object detectors (SS-OD) are based on pseudo-boxes, which need a sequence of post-processing with fine-tuned hyper-parameters. In this wo…