activity
20142024
most citedSynthetic Data and Artificial Neural Networks for Natural Scene Text Recognition

808 citations · 2.4k across the 23 of their papers we have counts for

collaborators

23 papers

cs.CV20243 cited

IM-3D: Iterative Multiview Diffusion and Reconstruction for High-Quality 3D Generation

Luke Melas-Kyriazi, Iro Laina, Christian Rupprecht +4

Most text-to-3D generators build upon off-the-shelf text-to-image models trained on billions of images. They use variants of Score Distillation Sampling (SDS), which is slow, somew…

cs.CV2024

Learning the 3D Fauna of the Web

Zizhang Li, Dor Litvak, Ruining Li +6

Learning 3D models of all animals on the Earth requires massively scaling up existing solutions. With this ultimate goal in mind, we develop 3D-Fauna, an approach that learns a pan…

cs.CV2023

HoloFusion: Towards Photo-realistic 3D Generative Modeling

Animesh Karnewar, Niloy J. Mitra, Andrea Vedaldi +1

Diffusion-based image generators can now produce high-quality and diverse samples, but their success has yet to fully translate to 3D generation: existing diffusion methods can eit…

cs.CV20232 cited

Online Clustered Codebook

Chuanxia Zheng, Andrea Vedaldi

Vector Quantisation (VQ) is experiencing a comeback in machine learning, where it is increasingly used in representation learning. However, optimizing the codevectors in existing V…

cs.CV2023

Replay: Multi-modal Multi-view Acted Videos for Casual Holography

Roman Shapovalov, Yanir Kleiman, Ignacio Rocco +6

We introduce Replay, a collection of multi-view, multi-modal videos of humans interacting socially. Each scene is filmed in high production quality, from different viewpoints with…

cs.CV20231 cited

DynamicStereo: Consistent Dynamic Depth from Stereo Videos

Nikita Karaev, Ignacio Rocco, Benjamin Graham +3

We consider the problem of reconstructing a dynamic scene observed from a stereo camera. Most existing methods for depth from stereo treat different stereo frames independently, le…