84 citations · 193 across the 18 of their papers we have counts for
34 papers
Long-term Video Frame Interpolation via Feature Propagation
Dawit Mureja Argaw, In So Kweon
Video frame interpolation (VFI) works generally predict intermediate frame(s) by first estimating the motion between inputs and then warping the inputs to the target time with the…
Audio-Visual Fusion Layers for Event Type Aware Video Recognition
Arda Senocak, Junsik Kim, Tae-Hyun Oh +3
Human brain is continuously inundated with the multisensory information and their complex interactions coming from the outside world at any given moment. Such information is automa…
Attentive and Contrastive Learning for Joint Depth and Motion Field Estimation
Seokju Lee, Francois Rameau, Fei Pan +1
Estimating the motion of the camera together with the 3D structure of the scene from a monocular vision system is a complex task that often relies on the so-called scene rigidity a…
Discover, Hallucinate, and Adapt: Open Compound Domain Adaptation for Semantic Segmentation
KwanYong Park, Sanghyun Woo, Inkyu Shin +1
Unsupervised domain adaptation (UDA) for semantic segmentation has been attracting attention recently, as it could be beneficial for various label-scarce real-world scenarios (e.g.…
Category-Level Metric Scale Object Shape and Pose Estimation
Taeyeop Lee, Byeong-Uk Lee, Myungchul Kim +1
Advances in deep learning recognition have led to accurate object detection with 2D images. However, these 2D perception methods are insufficient for complete 3D world information.…
VolumeFusion: Deep Depth Fusion for 3D Scene Reconstruction
Jaesung Choe, Sunghoon Im, Francois Rameau +2
To reconstruct a 3D scene from a set of calibrated views, traditional multi-view stereo techniques rely on two distinct stages: local depth maps computation and global depth maps f…