activity
20182021
most citedWord-level Deep Sign Language Recognition from Video: A New Large-scale Dataset and Methods Comparison

53 citations · 160 across the 16 of their papers we have counts for

collaborators
Showing cs.CVShow all

30 papers · 1 filter

cs.CV2021

PR-RRN: Pairwise-Regularized Residual-Recursive Networks for Non-rigid Structure-from-Motion

Haitian Zeng, Yuchao Dai, Xin Yu +2

We propose PR-RRN, a novel neural-network based method for Non-rigid Structure-from-Motion (NRSfM). PR-RRN consists of Residual-Recursive Networks (RRN) and two extra regularizatio…

cs.CV20212 cited

VidFace: A Full-Transformer Solver for Video FaceHallucination with Unaligned Tiny Snapshots

Yuan Gan, Yawei Luo, Xin Yu +2

In this paper, we investigate the task of hallucinating an authentic high-resolution (HR) human face from multiple low-resolution (LR) video snapshots. We propose a pure transforme…

cs.CV202136 cited

VTNet: Visual Transformer Network for Object Goal Navigation

Heming Du, Xin Yu, Liang Zheng

Object goal navigation aims to steer an agent towards a target object based on observations of the agent. It is of pivotal importance to design effective visual representations of…

cs.CV2021

Write-a-speaker: Text-based Emotional and Rhythmic Talking-head Generation

Lincheng Li, Suzhen Wang, Zhimeng Zhang +4

In this paper, we propose a novel text-based talking-head video generation framework that synthesizes high-fidelity facial expressions and head motions in accordance with contextua…

cs.CV2021

Blind Motion Deblurring Super-Resolution: When Dynamic Spatio-Temporal Learning Meets Static Image Understanding

Wenjia Niu, Kaihao Zhang, Wenhan Luo +1

Single-image super-resolution (SR) and multi-frame SR are two ways to super resolve low-resolution images. Single-Image SR generally handles each image independently, but ignores t…

cs.CV20212 cited

DSC-PoseNet: Learning 6DoF Object Pose Estimation via Dual-scale Consistency

Zongxin Yang, Xin Yu, Yi Yang

Compared to 2D object bounding-box labeling, it is very difficult for humans to annotate 3D object poses, especially when depth images of scenes are unavailable. This paper investi…