Showing cs.CVShow all
2 papers · 1 filter
cs.CV2024★ 1 cited
AddressCLIP: Empowering Vision-Language Models for City-wide Image Address Localization
Shixiong Xu, Chenghao Zhang, Lubin Fan +3
In this study, we introduce a new problem raised by social media and photojournalism, named Image Address Localization (IAL), which aims to predict the readable textual address whe…
cs.CV2024
Reusable Architecture Growth for Continual Stereo Matching
Chenghao Zhang, Gaofeng Meng, Bin Fan +4
The remarkable performance of recent stereo depth estimation models benefits from the successful use of convolutional neural networks to regress dense disparity. Akin to most tasks…