2 papers
cs.CV2025
Beyond Pixels: A Training-Free, Text-to-Text Framework for Remote Sensing Image Retrieval
J. Xiao, Y. Guo, X. Zi +3
Semantic retrieval of remote sensing (RS) images is a critical task fundamentally challenged by the \textquote{semantic gap}, the discrepancy between a model's low-level visual fea…
cs.RO2025
Large Language Models and 3D Vision for Intelligent Robotic Perception and Autonomy
Vinit Mehta, Charu Sharma, Karthick Thiyagarajan
With the rapid advancement of artificial intelligence and robotics, the integration of Large Language Models (LLMs) with 3D vision is emerging as a transformative approach to enhan…