260 citations · 280 across the 3 of their papers we have counts for
10 papers
MAAS: Multi-modal Assignation for Active Speaker Detection
Juan León-Alcázar, Fabian Caba Heilbron, Ali Thabet +1
Active speaker detection requires a solid integration of multi-modal cues. While individual modalities can approximate a solution, accurate predictions can only be achieved by expl…
SALA: Soft Assignment Local Aggregation for Parameter Efficient 3D Semantic Segmentation
Hani Itani, Silvio Giancola, Ali Thabet +1
In this work, we focus on designing a point local aggregation function that yields parameter efficient networks for 3D point cloud semantic segmentation. We explore the idea of usi…
MVTN: Multi-View Transformation Network for 3D Shape Recognition
Abdullah Hamdi, Silvio Giancola, Bernard Ghanem
Multi-view projection methods have demonstrated their ability to reach state-of-the-art performance on 3D shape recognition. Those methods learn different ways to aggregate informa…
LC-NAS: Latency Constrained Neural Architecture Search for Point Cloud Networks
Guohao Li, Mengmeng Xu, Silvio Giancola +2
Point cloud architecture design has become a crucial problem for 3D deep learning. Several efforts exist to manually design architectures with high accuracy in point cloud tasks su…
DeeperGCN: All You Need to Train Deeper GCNs
Guohao Li, Chenxin Xiong, Ali Thabet +1
Graph Convolutional Networks (GCNs) have been drawing significant attention with the power of representation learning on graphs. Unlike Convolutional Neural Networks (CNNs), which…
Gabor Layers Enhance Network Robustness
Juan C. Pérez, Motasem Alfarra, Guillaume Jeanneret +4
We revisit the benefits of merging classical vision concepts with deep learning models. In particular, we explore the effect on robustness against adversarial attacks of replacing…