2 papers
eess.AS2026
Cloned Voices, Real Consequences: Evaluating Bias in Political Deepfake Detection for Electoral Integrity in Brazil
Lucas Rafael Stefanel Gris, Daniel Casanova, Frederico Santos De Oliveira +4
Recent advances in generative artificial intelligence have made it easier to fabricate statements and amplify political disinformation during elections. We introduce ParlaSpoof-BR,…
cs.CL2026
Conv-to-Bench: Evaluating Language Models Via User-Assistant Dialogues In Code Tasks
Victor M. dos Santos, Andre C. Castro, Samuel L. de S. Toledo +5
The rapid advancement of Large Language Models (LLMs) has outpaced the scalability of traditional evaluation benchmarks, which remain heavily dependent on labor-intensive expert cu…