2 papers
cs.CV2024
HA-FGOVD: Highlighting Fine-grained Attributes via Explicit Linear Composition for Open-Vocabulary Object Detection
Yuqi Ma, Mengyin Liu, Chao Zhu +1
Open-vocabulary object detection (OVD) models are considered to be Large Multi-modal Models (LMM), due to their extensive training data and a large number of parameters. Mainstream…
cs.CL2024
SoftQE: Learned Representations of Queries Expanded by LLMs
Varad Pimpalkhute, John Heyer, Xusen Yin +1
We investigate the integration of Large Language Models (LLMs) into query encoders to improve dense retrieval without increasing latency and cost, by circumventing the dependency o…