papers

Publications (16)

cs.CV2018

Supervision-by-Registration: An Unsupervised Approach to Improve the Precision of Facial Landmark Detectors

Xuanyi Dong, Shoou-I Yu, Xinshuo Weng +3

In this paper, we present supervision-by-registration, an unsupervised approach to improve the precision of facial landmark detectors on both images and video. Our key observation…

cs.CV2015

The Best of Both Worlds: Combining Data-independent and Data-driven Approaches for Action Recognition

Zhenzhong Lan, Dezhong Yao, Ming Lin +2

Motivated by the success of data-driven convolutional neural networks (CNNs) in object recognition on static images, researchers are working hard towards developing CNN equivalents…

cs.CV2025

Relightable Full-Body Gaussian Codec Avatars

Shaofei Wang, Tomas Simon, Igor Santesteban +15

We propose Relightable Full-Body Gaussian Codec Avatars, a new approach for modeling relightable full-body avatars with fine-grained details including face and hands. The unique ch…

cs.CV2021

Supervision by Registration and Triangulation for Landmark Detection

Xuanyi Dong, Yi Yang, Shih-En Wei +3

We present Supervision by Registration and Triangulation (SRT), an unsupervised approach that utilizes unlabeled multi-view video to improve the accuracy and precision of landmark…

cs.CV2025

ATLAS: Decoupling Skeletal and Shape Parameters for Expressive Parametric Human Modeling

Jinhyung Park, Javier Romero, Shunsuke Saito +7

Parametric body models offer expressive 3D representation of humans across a wide range of poses, shapes, and facial expressions, typically derived by learning a basis over registe…

cs.IR2016

Strategies for Searching Video Content with Text Queries or Video Examples

Shoou-I Yu, Yi Yang, Zhongwen Xu +13

The large number of user-generated videos uploaded on to the Internet everyday has led to many commercial video search engines, which mainly rely on text metadata for search. Howev…

cs.CV2025

Generative Modeling of Shape-Dependent Self-Contact Human Poses

Takehiko Ohkawa, Jihyun Lee, Shunsuke Saito +7

One can hardly model self-contact of human poses without considering underlying body shapes. For example, the pose of rubbing a belly for a person with a low BMI leads to penetrati…

cs.CV2017

Long-Term Identity-Aware Multi-Person Tracking for Surveillance Video Summarization

Shoou-I Yu, Yi Yang, Xuanchong Li +1

Multi-person tracking plays a critical role in the analysis of surveillance video. However, most existing work focus on shorter-term (e.g. minute-long or hour-long) video sequences…

cs.CV2015

Handcrafted Local Features are Convolutional Neural Networks

Zhenzhong Lan, Shoou-I Yu, Ming Lin +2

Image and video classification research has made great progress through the development of handcrafted local features and learning based features. These two architectures were prop…

cs.CV2019

Self-Supervised Adaptation of High-Fidelity Face Models for Monocular Performance Tracking

Jae Shin Yoon, Takaaki Shiratori, Shoou-I Yu +1

Improvements in data-capture and face modeling techniques have enabled us to create high-fidelity realistic face models. However, driving these realistic face models requires speci…

cs.CV2021

CodedStereo: Learned Phase Masks for Large Depth-of-field Stereo

Shiyu Tan, Yicheng Wu, Shoou-I Yu +1

Conventional stereo suffers from a fundamental trade-off between imaging volume and signal-to-noise ratio (SNR) -- due to the conflicting impact of aperture size on both these vari…

cs.CV2020

Epipolar Transformers

Yihui He, Rui Yan, Katerina Fragkiadaki +1

A common approach to localize 3D human joints in a synchronized and calibrated multi-view setup consists of two-steps: (1) apply a 2D detector separately on each view to localize j…

cs.CV2023

Multiface: A Dataset for Neural Face Rendering

Cheng-hsin Wuu, Ningyuan Zheng, Scott Ardisson +28

Photorealistic avatars of human faces have come a long way in recent years, yet research along this area is limited by a lack of publicly available, high-quality datasets covering…

cs.CV2024

URHand: Universal Relightable Hands

Zhaoxi Chen, Gyeongsik Moon, Kaiwen Guo +20

Existing photorealistic relightable hand models require extensive identity-specific observations in different views, poses, and illuminations, and face challenges in generalizing t…

cs.GR2025

MHR: Momentum Human Rig

Aaron Ferguson, Ahmed A. A. Osman, Berta Bescos +44

We present MHR, a parametric human body model that combines the decoupled skeleton/shape paradigm of ATLAS with a flexible, modern rig and pose corrective system inspired by the Mo…

cs.CV2015

Improving Human Activity Recognition Through Ranking and Re-ranking

Zhenzhong Lan, Shoou-I Yu, Alexander G. Hauptmann

We propose two well-motivated ranking-based methods to enhance the performance of current state-of-the-art human activity recognition systems. First, as an improvement over the cla…