5 papers · 1 filter
Energy-Regularized Spatial Masking: A Novel Approach to Enhancing Robustness and Interpretability in Vision Models
Tom Devynck, Bilal Faye, Djamel Bouchaffra +3
Deep convolutional neural networks achieve remarkable performance by exhaustively processing dense spatial feature maps, yet this brute-force strategy introduces significant comput…
Lightweight Modular Parameter-Efficient Tuning for Open-Vocabulary Object Detection
Bilal Faye, Hanane Azzag, Mustapha Lebbah
Open-vocabulary object detection (OVD) extends recognition beyond fixed taxonomies by aligning visual and textual features, as in MDETR, GLIP, or RegionCLIP. While effective, these…
MB-ORES: A Multi-Branch Object Reasoner for Visual Grounding in Remote Sensing
Karim Radouane, Hanane Azzag, Mustapha lebbah
We propose a unified framework that integrates object detection (OD) and visual grounding (VG) for remote sensing (RS) imagery. To support conventional OD and establish an intuitiv…
OneEncoder: A Lightweight Framework for Progressive Alignment of Modalities
Bilal Faye, Hanane Azzag, Mustapha Lebbah
Cross-modal alignment Learning integrates information from different modalities like text, image, audio and video to create unified models. This approach develops shared representa…
Adaptative Context Normalization: A Boost for Deep Learning in Image Processing
Bilal Faye, Hanane Azzag, Mustapha Lebbah +1
Deep Neural network learning for image processing faces major challenges related to changes in distribution across layers, which disrupt model convergence and performance. Activati…