Showing cs.CVShow all
2 papers · 1 filter
cs.CV2026
Training-Free Open-Vocabulary Visual Grounding for Remote Sensing Images and Videos
Ke Li, Di Wang, Yongshan Zhu +5
Remote sensing visual grounding (RSVG) aims to localize a referred target in a remote sensing image or video according to a natural language expression. Existing RSVG methods usual…
cs.CV2024
TalkMosaic: Interactive PhotoMosaic with Multi-modal LLM Q&A Interactions
Kevin Li, Fulu Li
We use images of cars of a wide range of varieties to compose an image of an animal such as a bird or a lion for the theme of environmental protection to maximize the information a…