works on

From the 1 of 9 linked papers with an AI index.

most citedNAF: Zero-Shot Feature Upsampling via Neighborhood Attention Filtering

1 citations · 1 across the 8 of their papers we have counts for

collaborators
Showing cs.CVShow all

8 papers · 1 filter

cs.CV2026

Pictura: Perspective-View Self-Play at Scale for Driving

Yuan Yin, Elias Ramzi, Marc Lafon +8

The paper presents Pictura, a GPU‑accelerated multi‑agent driving simulator that renders each vehicle's egocentric camera view, enabling large‑scale self‑play training of driving p…

cs.CV2026

EditSSC: Toward Editable Semantic Occupancy Scenes with Unconditional Diffusion Models

Fatima Balde, Raoul de Charette, Alexandre Boulch

3D semantic scene generation is crucial for autonomous driving applications, yet most methods rely on complex 3D-specific architectures such as triplane encoders and adapted diffus…

cs.CV2026

Exploring Easy Boosts for Lidar Semantic Scene Completion

Tetiana Martyniuk, Jonathan Seele, Alexandre Boulch +3

This paper investigates "free lunch" strategies to boost the performance of lidar semantic scene completion (SSC) without requiring complex architectural redesigns. We first demons…

cs.CV2026

Vanilla ViT for Automotive Point Cloud Semantic Segmentation

Gilles Puy, Nermin Samet, Alexandre Boulch +3

Plain Transformers have become the de-facto architecture for processing text, audio, image, and video, offering a unified backbone for multimodal learning. However, state-of-the-ar…

cs.CV2026

SuperQuadricOcc: Real-Time Self-Supervised Semantic Occupancy Estimation with Superquadric Volume Rendering

Seamie Hayes, Alexandre Boulch, Andrei Bursuc +4

Self-supervision for semantic occupancy estimation is appealing as it removes the labour-intensive manual annotation, thus allowing one to scale to larger autonomous driving datase…

cs.CV2026

Driving on Registers

Ellington Kirby, Alexandre Boulch, Yihong Xu +11

We present DrivoR, a simple and efficient transformer-based architecture for end-to-end autonomous driving. Our approach builds on pretrained Vision Transformers (ViTs) and introdu…