13 citations · 14 across the 2 of their papers we have counts for
Showing cs.CVShow all
3 papers · 1 filter
cs.CV2025★ 1 cited
A Simple Aerial Detection Baseline of Multimodal Language Models
Qingyun Li, Yushi Chen, Xinya Shu +4
The multimodal language models (MLMs) based on generative pre-trained Transformer are considered powerful candidates for unifying various domains and tasks. MLMs developed for remo…
cs.CV2023★ 13 cited
The All-Seeing Project: Towards Panoptic Visual Recognition and Understanding of the Open World
Weiyun Wang, Min Shi, Qingyun Li +11
We present the All-Seeing (AS) project: a large-scale data and model for recognizing and understanding everything in the open world. Using a scalable data engine that incorporates…
cs.CV2023
ARS-DETR: Aspect Ratio-Sensitive Detection Transformer for Aerial Oriented Object Detection
Ying Zeng, Yushi Chen, Xue Yang +2
Existing oriented object detection methods commonly use metric AP to measure the performance of the model. We argue that AP is inherently unsuitable for oriented obje…