30 citations · 36 across the 3 of their papers we have counts for
3 papers
cs.CV2023★ 1 cited
Visual In-Context Prompting
Feng Li, Qing Jiang, Hao Zhang +9
In-context prompting in large language models (LLMs) has become a prevalent approach to improve zero-shot capabilities, but this idea is less explored in the vision domain. Existin…
cs.CV2023★ 5 cited
T-Rex: Counting by Visual Prompting
Qing Jiang, Feng Li, Tianhe Ren +4
We introduce T-Rex, an interactive object counting model designed to first detect and then count any objects. We formulate object counting as an open-set object detection task with…
cs.CV2022★ 30 cited
Vision-Language Intelligence: Tasks, Representation Learning, and Large Models
Feng Li, Hao Zhang, Yi-Fan Zhang +5
This paper presents a comprehensive survey of vision-language (VL) intelligence from the perspective of time. This survey is inspired by the remarkable progress in both computer vi…