2 citations · 3 across the 10 of their papers we have counts for
5 papers · 1 filter
Energy-Regularized Spatial Masking: A Novel Approach to Enhancing Robustness and Interpretability in Vision Models
Tom Devynck, Bilal Faye, Djamel Bouchaffra +3
Deep convolutional neural networks achieve remarkable performance by exhaustively processing dense spatial feature maps, yet this brute-force strategy introduces significant comput…
MB-ORES: A Multi-Branch Object Reasoner for Visual Grounding in Remote Sensing
Karim Radouane, Hanane Azzag, Mustapha lebbah
We propose a unified framework that integrates object detection (OD) and visual grounding (VG) for remote sensing (RS) imagery. To support conventional OD and establish an intuitiv…
OneEncoder: A Lightweight Framework for Progressive Alignment of Modalities
Bilal Faye, Hanane Azzag, Mustapha Lebbah
Cross-modal alignment Learning integrates information from different modalities like text, image, audio and video to create unified models. This approach develops shared representa…
Lightweight Modular Parameter-Efficient Tuning for Open-Vocabulary Object Detection
Bilal Faye, Hanane Azzag, Mustapha Lebbah
Open-vocabulary object detection (OVD) extends recognition beyond fixed taxonomies by aligning visual and textual features, as in MDETR, GLIP, or RegionCLIP. While effective, these…
Context Normalization Layer with Applications
Bilal Faye, Mohamed-Djallel Dilmi, Hanane Azzag +2
Normalization is a pre-processing step that converts the data into a more usable representation. As part of the deep neural networks (DNNs), the batch normalization (BN) technique…