most citedCounterfeit Answers: Adversarial Forgery against OCR-Free Document Visual Question Answering

1 citations · 1 across the 2 of their papers we have counts for

collaborators

20 papers

cs.CV2026

Rethinking Expert Training for Model Merging with Prompt Learning

Christos Georgakilas, Aniello Panariello, Samir El Karrat Moreno +3

Model merging aims to combine multiple domain-specialized experts trained from a shared foundation model into a single multi-task model. Existing approaches largely focus on improv…

cs.CV20261 cited

Counterfeit Answers: Adversarial Forgery against OCR-Free Document Visual Question Answering

Marco Pintore, Maura Pintor, Dimosthenis Karatzas +1

Document Visual Question Answering (DocVQA) enables end-to-end reasoning grounded on information present in a document input. While recent models have shown impressive capabilities…

cs.CV2026

SMI: Efficient Self-Supervised Learning via Mutual-Information-Inspired Dependency Optimization

Pritam Mishra, Coloma Ballester, Dimosthenis Karatzas

Self-supervised learning (SSL) has achieved remarkable representation learning performance, but many existing methods rely on large batch sizes, memory banks, momentum encoders, or…

cs.CV2026

TRIM: A Self-Supervised Video Summarization Framework Maximizing Temporal Relative Information and Representativeness

Pritam Mishra, Coloma Ballester, Dimosthenis Karatzas

The increasing ubiquity of video content and the corresponding demand for efficient access to meaningful information have elevated video summarization and video highlights as a vit…

cs.AI2026

Learning Quantifiable Visual Explanations Without Ground-Truth

Amritpal Singh, Andrey Barsky, Mohamed Ali Souibgui +2

Explainable AI (XAI) techniques are increasingly important for the validation and responsible use of modern deep learning models, but are difficult to evaluate due to the lack of g…

cs.CV2026

TRIMMER: A New Paradigm for Video Summarization through Self-Supervised Reinforcement Learning

Pritam Mishra, Coloma Ballester, Dimosthenis Karatzas

The rapid growth of video content across domains such as surveillance, education, and social media has made efficient content understanding increasingly critical. Video summarizati…