4 papers
CARTE: A Benchmark for Mapping Language Model Knowledge Across France
Sarah Almeida Carneiro, Christos Xypolopoulos, Xiao Fei +2
We introduce CARTE 1 (Culturally Anchored Regional-Territorial Evaluation), a multiplechoice benchmark for evaluating the ability of large language models (LLMs) to perform fine-gr…
GreekMMLU: A Native-Sourced Multitask Benchmark for Evaluating Language Models in Greek
Yang Zhang, Mersin Konomi, Christos Xypolopoulos +6
Large Language Models (LLMs) are commonly trained on multilingual corpora that include Greek, yet reliable evaluation benchmarks for Greek-particularly those based on authentic, na…
Graph Linearization Methods for Reasoning on Graphs with Large Language Models
Christos Xypolopoulos, Guokan Shang, Xiao Fei +6
Large language models have evolved to process multiple modalities beyond text, such as images and audio, which motivates us to explore how to effectively leverage them for graph re…
Bias in the Mirror: Are LLMs opinions robust to their own adversarial attacks ?
Virgile Rennard, Christos Xypolopoulos, Michalis Vazirgiannis
Large language models (LLMs) inherit biases from their training data and alignment processes, influencing their responses in subtle ways. While many studies have examined these bia…