10 citations · 10 across the 1 of their papers we have counts for
2 papers
cs.CV2025
Beyond Pixels: A Training-Free, Text-to-Text Framework for Remote Sensing Image Retrieval
J. Xiao, Y. Guo, X. Zi +3
Semantic retrieval of remote sensing (RS) images is a critical task fundamentally challenged by the \textquote{semantic gap}, the discrepancy between a model's low-level visual fea…
cs.RO2025★ 10 cited
Large Language Models and 3D Vision for Intelligent Robotic Perception and Autonomy
Vinit Mehta, Charu Sharma, Karthick Thiyagarajan
With the rapid advancement of artificial intelligence and robotics, the integration of Large Language Models (LLMs) with 3D vision is emerging as a transformative approach to enhan…