activity
20202026
most citedCroCo: Self-Supervised Pre-training for 3D Vision Tasks by Cross-View Completion

16 citations · 22 across the 9 of their papers we have counts for

collaborators
Showing cs.CVShow all

12 papers · 1 filter

cs.CV2026

Sparse auto-regressive modeling for scene generation from multi-view images

Thomas Lucas, Maxime Pietrantoni, Philippe Weinzaepfel +4

Generating complete 3D scenes from sparse, unconstrained views is a fundamental challenge in 3D vision which requires reasoning beyond observed content while remaining computationa…

cs.CV2023★ 3 cited

DUSt3R: Geometric 3D Vision Made Easy

Shuzhe Wang, Vincent Leroy, Yohann Cabon +2

Multi-view stereo reconstruction (MVS) in the wild requires to first estimate the camera parameters e.g. intrinsic and extrinsic parameters. These are usually tedious and cumbersom…

cs.CV2023

Cross-view and Cross-pose Completion for 3D Human Understanding

Matthieu Armando, Salma Galaaoui, Fabien Baradel +5

Human perception and understanding is a major domain of computer vision which, like many other vision subdomains recently, stands to gain from the use of large models pre-trained o…

cs.CV2023

Win-Win: Training High-Resolution Vision Transformers from Two Windows

Vincent Leroy, Jerome Revaud, Thomas Lucas +1

Transformers have become the standard in state-of-the-art vision architectures, achieving impressive performance on both image-level and dense pixelwise tasks. However, training vi…

cs.CV2023

SHOWMe: Benchmarking Object-agnostic Hand-Object 3D Reconstruction

Anilkumar Swamy, Vincent Leroy, Philippe Weinzaepfel +6

Recent hand-object interaction datasets show limited real object variability and rely on fitting the MANO parametric model to obtain groundtruth hand shapes. To go beyond these lim…

cs.CV2023★ 1 cited

4DHumanOutfit: a multi-subject 4D dataset of human motion sequences in varying outfits exhibiting large displacements

Matthieu Armando, Laurence Boissieux, Edmond Boyer +11

This work presents 4DHumanOutfit, a new dataset of densely sampled spatio-temporal 4D human motion data of different actors, outfits and motions. The dataset is designed to contain…