Showing cs.CVShow all
2 papers · 1 filter
cs.CV2025
MoniRefer: A Real-world Large-scale Multi-modal Dataset based on Roadside Infrastructure for 3D Visual Grounding
Panquan Yang, Junfei Huang, Zongzhangbao Yin +9
3D visual grounding aims to localize the object in 3D point cloud scenes that semantically corresponds to given natural language sentences. It is very critical for roadside infrast…
cs.CV2019
Meta R-CNN : Towards General Solver for Instance-level Few-shot Learning
Xiaopeng Yan, Ziliang Chen, Anni Xu +3
Resembling the rapid learning capability of human, few-shot learning empowers vision systems to understand new concepts by training with few samples. Leading approaches derived fro…