Showing cs.CVShow all
2 papers · 1 filter
cs.CV2025
Language-Image Alignment with Fixed Text Encoders
Jingfeng Yang, Ziyang Wu, Yue Zhao +1
Currently, the most dominant approach to establishing language-image alignment is to pre-train text and image encoders jointly through contrastive learning, such as CLIP and its va…
cs.CV2020
LR-CNN: Local-aware Region CNN for Vehicle Detection in Aerial Imagery
Wentong Liao, Xiang Chen, Jingfeng Yang +4
State-of-the-art object detection approaches such as Fast/Faster R-CNN, SSD, or YOLO have difficulties detecting dense, small targets with arbitrary orientation in large aerial ima…