Automatic Generation of Dense Non-rigid Optical Flow
arXiv:1812.01946 · doi:10.1016/j.cviu.2021.103274
Abstract
There hardly exists any large-scale datasets with dense optical flow of non-rigid motion from real-world imagery as of today. The reason lies mainly in the required setup to derive ground truth optical flows: a series of images with known camera poses along its trajectory, and an accurate 3D model from a textured scene. Human annotation is not only too tedious for large databases, it can simply hardly contribute to accurate optical flow. To circumvent the need for manual annotation, we propose a framework to automatically generate optical flow from real-world videos. The method extracts and matches objects from video frames to compute initial constraints, and applies a deformation over the objects of interest to obtain dense optical flow fields. We propose several ways to augment the optical flow variations. Extensive experimental results show that training on our automatically generated optical flow outperforms methods that are trained on rigid synthetic data using FlowNet-S, LiteFlowNet, PWC-Net, and RAFT. Datasets and implementation of our optical flow generation framework are released at https://github.com/lhoangan/arap_flow
The paper is accepted for publication for Computer Vision and Image Understanding (CVIU)
References in corpus (14)
- Two-Stream Convolutional Networks for Action Recognition in Videos
- Depth Map Prediction from a Single Image using a Multi-Scale Deep Network
- FlowNet: Learning Optical Flow with Convolutional Networks
- What Makes Good Synthetic Training Data for Learning Disparity and Optical Flow Estimation?
- Neighbourhood Consensus Networks
- RAFT: Recurrent All-Pairs Field Transforms for Optical Flow
- Accurate Optical Flow via Direct Cost Volume Processing
- Optical Flow Guided Feature: A Fast and Robust Motion Representation for Video Action Recognition
- SipMask: Spatial Information Preservation for Fast Image and Video Instance Segmentation
- End-to-end Flow Correlation Tracking with Spatial-temporal Attention
- Learning Human Optical Flow
- DDFlow: Learning Optical Flow with Unlabeled Data Distillation
- SelFlow: Self-Supervised Learning of Optical Flow
- Im2Flow: Motion Hallucination from Static Images for Action Recognition