most citedRAVEN: A Dataset for Relational and Analogical Visual rEasoNing

31 citations · 53 across the 5 of their papers we have counts for

collaborators

5 papers

cs.CV2023

Self-Supervised 3D Scene Flow Estimation and Motion Prediction using Local Rigidity Prior

Ruibo Li, Chi Zhang, Zhe Wang +2

In this article, we investigate self-supervised 3D scene flow estimation and class-agnostic motion prediction on point clouds. A realistic scene can be well modeled as a collection…

cs.CV20234 cited

Metric3D: Towards Zero-shot Metric 3D Prediction from A Single Image

Wei Yin, Chi Zhang, Hao Chen +5

Reconstructing accurate 3D scenes from images is a long-standing vision task. Due to the ill-posedness of the single-image reconstruction problem, most well-established methods are…

cs.CL20235 cited

MEWL: Few-shot multimodal word learning with referential uncertainty

Guangyuan Jiang, Manjie Xu, Shiji Xin +4

Without explicit feedback, humans can rapidly learn the meaning of words. Children can acquire a new word after just a few passive exposures, a process known as fast mapping. This…

cs.CV202313 cited

StyleAvatar3D: Leveraging Image-Text Diffusion Models for High-Fidelity 3D Avatar Generation

Chi Zhang, Yiwen Chen, Yijun Fu +7

The recent advancements in image-text diffusion models have stimulated research interest in large-scale 3D generative models. Nevertheless, the limited availability of diverse 3D r…

cs.CV201931 cited

RAVEN: A Dataset for Relational and Analogical Visual rEasoNing

Chi Zhang, Feng Gao, Baoxiong Jia +2

Dramatic progress has been witnessed in basic vision tasks involving low-level perception, such as object recognition, detection, and tracking. Unfortunately, there is still an eno…