activity
20232026
most citedPuzzle Solving using Reasoning of Large Language Models: A Survey

9 citations · 14 across the 41 of their papers we have counts for

collaborators

43 papers

cs.CL2026

Benchmarking Gender Bias in Machine Translation Evaluation Metrics across Occupations

Orfeas Menis Mastromichalakis, Giorgos Filandrianos, Wafaa Mohammed +2

Gender bias remains a persistent concern in machine translation (MT), affecting both generated translations and their automatic evaluation. When a source text leaves a person's gen…

cs.CL2026

An Empirical Study of Counterfactual Self-Explanations in LLMs

Giannis Kalyvas, Giorgos Filandrianos, Orfeas Menis Mastromichalakis +2

Large language models can easily generate explanations for their own outputs, but such self-explanations are not necessarily faithful to the model's behavior. We study this issue t…

cs.AI2026

U-CECE: A Universal Multi-Resolution Framework for Conceptual Counterfactual Explanations

Angeliki Dimitriou, Nikolaos Chaidos, Maria Lymperaiou +2

As AI models grow more complex, explainability is essential for building trust, yet concept-based counterfactual methods still face a trade-off between expressivity and efficiency.…

cs.CL2026

AILS-NTUA at SemEval-2026 Task 8: Evaluating Multi-Turn RAG Conversations

Dimosthenis Athanasiou, Maria Lymperaiou, Giorgos Filandrianos +2

We present the AILS-NTUA system for SemEval-2026 Task 8 (MTRAGEval), addressing all three subtasks of multi-turn retrieval-augmented generation: passage retrieval (A), reference-gr…

cs.CL2026

AILS-NTUA at SemEval-2026 Task 3: Efficient Dimensional Aspect-Based Sentiment Analysis

Stavros Gazetas, Giorgos Filandrianos, Maria Lymperaiou +3

In this paper, we present AILS-NTUA system for Track-A of SemEval-2026 Task 3 on Dimensional Aspect-Based Sentiment Analysis (DimABSA), which encompasses three complementary proble…

cs.CL2026

AILS-NTUA at SemEval-2026 Task 10: Agentic LLMs for Psycholinguistic Marker Extraction and Conspiracy Endorsement Detection

Panagiotis Alexios Spanakis, Maria Lymperaiou, Giorgos Filandrianos +2

This paper presents a novel agentic LLM pipeline for SemEval-2026 Task 10 that jointly extracts psycholinguistic conspiracy markers and detects conspiracy endorsement. Unlike tradi…