3 papers
cs.AI2026
GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks
Rodrigo Ferreira Rodrigues, Karim Radouane, Jose G Moreno +1
In the context of geodata, existing Large Language Models have often been studied in a homogeneous setting, which has considerably limited insights into their generalization capabi…
cs.CL2026
Geo-Spatial Concept Probing of Large Language Models: Abstraction, Compositionality, and Grounding
Karim Radouane, Jose G Moreno, Lynda Tamine
Understanding concepts is fundamental to generalization. Despite their impressive performance on a wide range of tasks, Large Language Models (LLMs) still struggle with genuine con…
cs.CL2025
Jointly Generating and Attributing Answers using Logits of Document-Identifier Tokens
Lucas Albarede, Jose Moreno, Lynda Tamine +1
Despite their impressive performances, Large Language Models (LLMs) remain prone to hallucination, which critically undermines their trustworthiness. While most of the previous wor…