collaborators

8 papers

cs.CV2026

AMALIA-VL: A Native European Portuguese Open-Source Vision and Language Model

Diogo Glória-Silva, João Cardeira, Manuel Letras da Luz +8

Large Vision and Language Models (LVLMs) have advanced rapidly, yet European Portuguese (pt-PT) remains systematically underserved by existing open-source multimodal models, which…

cs.CV2026

PorTEXTO: A European Portuguese Benchmark for Visual Text Extraction

João Cardeira, Diogo Glória-Silva, Manuel Letras da Luz +4

European Portuguese (pt-PT) is largely absent from OCR benchmarks, which skew toward high-resource languages. The few benchmarks that cover pt-PT focus on historical artifacts and…

cs.CL2026

P3B3: A Multi-Turn Conversational Benchmark for Measuring European and Brazilian Portuguese Variety Bias in LLMs

Rafael Ferreira, Inês Vieira, Inês Calvo +6

As Large Language Models (LLMs) become embedded in everyday communication, capturing regional linguistic variation is essential for reliable and equitable language use. In Portugue…

cs.CV2026

VIGiA: Instructional Video Guidance via Dialogue Reasoning and Retrieval

Diogo Glória-Silva, David Semedo, João Maglhães

We introduce VIGiA, a novel multimodal dialogue model designed to understand and reason over complex, multi-step instructional video action plans. Unlike prior work which focuses m…

cs.CV2026

FineVAU: A Novel Human-Aligned Benchmark for Fine-Grained Video Anomaly Understanding

João Pereira, Vasco Lopes, João Neves +1

Video Anomaly Understanding (VAU) is a novel task focused on describing unusual occurrences in videos. Despite growing interest, the evaluation of VAU remains an open challenge. Ex…

cs.CV2025

Chain-of-Anomaly Thoughts with Large Vision-Language Models

Pedro Domingos, João Pereira, Vasco Lopes +2

Automated video surveillance with Large Vision-Language Models is limited by their inherent bias towards normality, often failing to detect crimes. While Chain-of-Thought reasoning…