13 citations · 30 across the 19 of their papers we have counts for
31 papers · 1 filter
ProactiveBench: Benchmarking Proactiveness in Multimodal Large Language Models
Thomas De Min, Subhankar Roy, Stéphane Lathuilière +2
Effective collaboration begins with knowing when to ask for help. For example, when trying to identify an occluded object, a human would ask someone to remove the obstruction. Can…
DiO: Distilling Masked Diffusion Models into One-step Generator
Yuanzhi Zhu, Xi Wang, Stéphane Lathuilière +1
Masked Diffusion Models (MDMs) have emerged as a powerful generative modeling technique. Despite their remarkable results, they typically suffer from slow inference with several st…
Ask and Remember: A Questions-Only Replay Strategy for Continual Visual Question Answering
Imad Eddine Marouf, Enzo Tartaglione, Stephane Lathuiliere +1
Continual Learning in Visual Question Answering (VQACL) requires models to acquire new visual-linguistic skills (plasticity) while preserving previously learned knowledge (stabilit…
Unlearning Personal Data from a Single Image
Thomas De Min, Massimiliano Mancini, Stéphane Lathuilière +2
Machine unlearning aims to erase data from a model as if the latter never saw them during training. While existing approaches unlearn information from complete or partial access to…
Collaborating Foundation Models for Domain Generalized Semantic Segmentation
Yasser Benigmim, Subhankar Roy, Slim Essid +2
Domain Generalized Semantic Segmentation (DGSS) deals with training a model on a labeled source domain with the aim of generalizing to unseen domains during inference. Existing DGS…
Mini but Mighty: Finetuning ViTs with Mini Adapters
Imad Eddine Marouf, Enzo Tartaglione, Stéphane Lathuilière
Vision Transformers (ViTs) have become one of the dominant architectures in computer vision, and pre-trained ViT models are commonly adapted to new tasks via fine-tuning. Recent wo…