Video Inpainting of Complex Scenes
arXiv:1503.05528 · doi:10.1137/140954933
Abstract
We propose an automatic video inpainting algorithm which relies on the optimisation of a global, patch-based functional. Our algorithm is able to deal with a variety of challenging situations which naturally arise in video inpainting, such as the correct reconstruction of dynamic textures, multiple moving objects and moving background. Furthermore, we achieve this in an order of magnitude less execution time with respect to the state-of-the-art. We are also able to achieve good quality results on high definition videos. Finally, we provide specific algorithmic details to make implementation of our algorithm as easy as possible. The resulting algorithm requires no segmentation or manual input other than the definition of the inpainting mask, and can deal with a wider variety of situations than is handled by previous work. 1. Introduction. Advanced image and video editing techniques are increasingly common in the image processing and computer vision world, and are also starting to be used in media entertainment. One common and difficult task closely linked to the world of video editing is image and video " inpainting ". Generally speaking, this is the task of replacing the content of an image or video with some other content which is visually pleasing. This subject has been extensively studied in the case of images, to such an extent that commercial image inpainting products destined for the general public are available, such as Photoshop's " Content Aware fill " [1]. However, while some impressive results have been obtained in the case of videos, the subject has been studied far less extensively than image inpainting. This relative lack of research can largely be attributed to high time complexity due to the added temporal dimension. Indeed, it has only very recently become possible to produce good quality inpainting results on high definition videos, and this only in a semi-automatic manner. Nevertheless, high-quality video inpainting has many important and useful applications such as film restoration, professional post-production in cinema and video editing for personal use. For this reason, we believe that an automatic, generic video inpainting algorithm would be extremely useful for both academic and professional communities.
Cited by in corpus (37)
- Generative Image Inpainting with Contextual Attention
- Learnable Gated Temporal Shift Module for Deep Video Inpainting
- Decoupled Spatial-Temporal Transformer for Video Inpainting
- Structural inpainting
- Deep Face Video Inpainting via UV Mapping
- Iterative training of neural networks for intra prediction
- Unveiling the invisible - mathematical methods for restoring and interpreting illuminated manuscripts
- Learning Joint Spatial-Temporal Transformations for Video Inpainting
- An Internal Learning Approach to Video Inpainting
- Exploiting Optical Flow Guidance for Transformer-Based Video Inpainting
- DeViT: Deformed Vision Transformers in Video Inpainting
- Video Inpainting by Jointly Learning Temporal Structure and Spatial Details
- First image then video: A two-stage network for spatiotemporal video denoising
- Free-form Video Inpainting with 3D Gated Convolution and Temporal PatchGAN
- FuseFormer: Fusing Fine-Grained Information in Transformers for Video Inpainting
- VORNet: Spatio-temporally Consistent Video Inpainting for Object Removal
- Deep Flow-Guided Video Inpainting
- TransFill: Reference-guided Image Inpainting by Merging Multiple Color and Spatial Transformations
- ChaLearn Looking at People: Inpainting and Denoising challenges
- Multi-View Frame Reconstruction with Conditional GAN
- A Temporally-Aware Interpolation Network for Video Frame Inpainting
- AutoRemover: Automatic Object Removal for Autonomous Driving Videos
- Short-Term and Long-Term Context Aggregation Network for Video Inpainting
- Deep Blind Video Decaptioning by Temporal Aggregation and Recurrence
- Progressive Temporal Feature Alignment Network for Video Inpainting
- DVI: Depth Guided Video Inpainting for Autonomous Driving
- Restore from Restored: Single-image Inpainting
- A nonlocal feature-driven exemplar-based approach for image inpainting
- Patch-based field-of-view matching in multi-modal images for electroporation-based ablations
- Omnimatte: Associating Objects and Their Effects in Video
- Flow-edge Guided Video Completion
- Dim the Lights! -- Low-Rank Prior Temporal Data for Specular-Free Video Recovery
- Multi-View Inpainting for RGB-D Sequence
- Infusion: internal diffusion for inpainting of dynamic textures and complex motion
- Generating Videos of Zero-Shot Compositions of Actions and Objects
- Layered Neural Rendering for Retiming People in Video
- Guidefill: GPU Accelerated, Artist Guided Geometric Inpainting for 3D Conversion