7.1k citations
- Centre National de la Recherche ScientifiqueFR406 papers
- Centre de Recherche en Informatique, Signal et Automatique de LilleFR188 papers
- École Centrale de LilleFR162 papers
- Sorbonne UniversitéFR158 papers
- Laboratoire Paul PainlevéFR141 papers
- Université Paris CitéFR118 papers
- Université Paris Sciences et LettresFR117 papers
- Centre de Recherche en InformatiqueFR116 papers
- Commissariat à l'Énergie Atomique et aux Énergies AlternativesFR102 papers
- CEA Paris-SaclayFR94 papers
- Université Paris-SaclayFR94 papers
- Laboratoire de Physique des PlasmasFR92 papers
22 papers · 1 filter
Privacy-Preserving Person Re-Identification from Temporal Sequences with Transformer and Hungarian Optimization
Raphaël Delécluse, Hazem Wannous, Laurent Guimas
Person re-identification (Re-ID) is a crucial task in surveillance and human behavior analysis, often used in public spaces such as transport hubs. Traditional RGB-based Re-ID meth…
DinoLizer: Separating VAE and Diffusion Artifacts in Generative Inpainting Localization
Minh Thong Doi, Vincent Itier, Jan Butora +2
We introduce DinoLizer, a DINOv2-based localizer of manipulated areas in generative inpainting. The model is trained to focus on semantically altered regions by treating reconstruc…
RAVID: Retrieval-Augmented Visual Detection: A Knowledge-Driven Approach for AI-Generated Image Identification
Mamadou Keita, Wassim Hamidouche, Hessen Bougueffa Eutamene +2
In this paper, we introduce RAVID, the first framework for AI-generated image detection that leverages visual retrieval-augmented generation (RAG). While RAG methods have shown pro…
LoLA-SpecViT: Local Attention SwiGLU Vision Transformer with LoRA for Hyperspectral Imaging
Fadi Abdeladhim Zidi, Djamel Eddine Boukhari, Abdellah Zakaria Sellam +4
Hyperspectral image classification remains a challenging task due to the high dimensionality of spectral data, significant inter-band redundancy, and the limited availability of an…
PE-CLIP: A Parameter-Efficient Fine-Tuning of Vision Language Models for Dynamic Facial Expression Recognition
Ibtissam Saadi, Abdenour Hadid, Douglas W. Cunningham +2
Vision-Language Models (VLMs) like CLIP offer promising solutions for Dynamic Facial Expression Recognition (DFER) but face challenges such as inefficient full fine-tuning, high co…
Online hand gesture recognition using Continual Graph Transformers
Rim Slama, Wael Rabah, Hazem Wannous
Online continuous action recognition has emerged as a critical research area due to its practical implications in real-world applications, such as human-computer interaction, healt…