activity
20192025
most citedLocality Aware Appearance Metric for Multi-Target Multi-Camera Tracking

16 citations · 36 across the 5 of their papers we have counts for

collaborators

10 papers

cs.CV2025

Extreme Amodal Face Detection

Changlin Song, Yunzhong Hou, Michael Randall Barnes +2

Extreme amodal detection is the task of inferring the 2D location of objects that are not fully visible in the input image but are visible within an expanded field-of-view. This di…

cs.CV2025

Effective Training Data Synthesis for Improving MLLM Chart Understanding

Yuwei Yang, Zeyu Zhang, Yunzhong Hou +5

Being able to effectively read scientific plots, or chart understanding, is a central part toward building effective agents for science. However, existing multimodal large language…

cs.CV2025

REPA-E: Unlocking VAE for End-to-End Tuning with Latent Diffusion Transformers

Xingjian Leng, Jaskirat Singh, Yunzhong Hou +3

In this paper we tackle a fundamental question: "Can we train latent diffusion models together with the variational auto-encoder (VAE) tokenizer in an end-to-end manner?" Tradition…

cs.CV2024

Learning Camera Movement Control from Real-World Drone Videos

Yunzhong Hou, Liang Zheng, Philip Torr

This study seeks to automate camera movement control for filming existing subjects into attractive videos, contrasting with the creation of non-existent content by directly generat…

cs.CV2021

Ranking Models in Unlabeled New Environments

Xiaoxiao Sun, Yunzhong Hou, Weijian Deng +2

Consider a scenario where we are supplied with a number of ready-to-use models trained on a certain source domain and hope to directly apply the most appropriate ones to different…

cs.CV20213 cited

Memory-Free Generative Replay For Class-Incremental Learning

Xiaomeng Xin, Yiran Zhong, Yunzhong Hou +2

Regularization-based methods are beneficial to alleviate the catastrophic forgetting problem in class-incremental learning. With the absence of old task images, they often assume t…