3 papers
cs.CV2026
Training-Free Open-Vocabulary Visual Grounding for Remote Sensing Images and Videos
Ke Li, Di Wang, Yongshan Zhu +5
Remote sensing visual grounding (RSVG) aims to localize a referred target in a remote sensing image or video according to a natural language expression. Existing RSVG methods usual…
cs.CV2024
TalkMosaic: Interactive PhotoMosaic with Multi-modal LLM Q&A Interactions
Kevin Li, Fulu Li
We use images of cars of a wide range of varieties to compose an image of an animal such as a bird or a lion for the theme of environmental protection to maximize the information a…
cs.AI2024
Analysis on Riemann Hypothesis with Cross Entropy Optimization and Reasoning
Kevin Li, Fulu Li
In this paper, we present a novel framework for the analysis of Riemann Hypothesis [27], which is composed of three key components: a) probabilistic modeling with cross entropy opt…