1 citations · 1 across the 3 of their papers we have counts for
3 papers
cs.CV2024
RSA: Resolving Scale Ambiguities in Monocular Depth Estimators through Language Descriptions
Ziyao Zeng, Yangchao Wu, Hyoungseob Park +6
We propose a method for metric-scale monocular depth estimation. Inferring depth from a single image is an ill-posed problem due to the loss of scale from perspective projection du…
q-bio.NC2024
NeuroBind: Towards Unified Multimodal Representations for Neural Signals
Fengyu Yang, Chao Feng, Daniel Wang +8
Understanding neural activity and information representation is crucial for advancing knowledge of brain function and cognition. Neural activity, measured through techniques like e…
cs.CV2022★ 1 cited
Can Language Understand Depth?
Renrui Zhang, Ziyao Zeng, Ziyu Guo +1
Besides image classification, Contrastive Language-Image Pre-training (CLIP) has accomplished extraordinary success for a wide range of vision tasks, including object-level and 3D…