activity
20242026
collaborators

6 papers

cs.CV2026

AMALIA-VL: A Native European Portuguese Open-Source Vision and Language Model

Diogo Glória-Silva, João Cardeira, Manuel Letras da Luz +8

Large Vision and Language Models (LVLMs) have advanced rapidly, yet European Portuguese (pt-PT) remains systematically underserved by existing open-source multimodal models, which…

cs.CV2026

PorTEXTO: A European Portuguese Benchmark for Visual Text Extraction

João Cardeira, Diogo Glória-Silva, Manuel Letras da Luz +4

European Portuguese (pt-PT) is largely absent from OCR benchmarks, which skew toward high-resource languages. The few benchmarks that cover pt-PT focus on historical artifacts and…

cs.CL2026

Global PIQA: Evaluating Commonsense Reasoning Across 100+ Languages and Cultures

Tyler A. Chang, Catherine Arnett, Abdelrahman Sadallah +377

To date, there exist almost no culturally-specific evaluation benchmarks for large language models (LLMs) that cover a large number of languages and cultures. In this paper, we pre…

cs.CR2025

RedTWIZ: Diverse LLM Red Teaming via Adaptive Attack Planning

Artur Horal, Daniel Pina, Henrique Paz +7

This paper presents the vision, scientific contributions, and technical details of RedTWIZ: an adaptive and diverse multi-turn red teaming framework, to audit the robustness of Lar…

cs.CL2024

Multi-trait User Simulation with Adaptive Decoding for Conversational Task Assistants

Rafael Ferreira, David Semedo, João Magalhães

Conversational systems must be robust to user interactions that naturally exhibit diverse conversational traits. Capturing and simulating these diverse traits coherently and effici…

cs.CV2024

Show and Guide: Instructional-Plan Grounded Vision and Language Model

Diogo Glória-Silva, David Semedo, João Magalhães

Guiding users through complex procedural plans is an inherently multimodal task in which having visually illustrated plan steps is crucial to deliver an effective plan guidance. Ho…