17 citations · 17 across the 1 of their papers we have counts for
3 papers
cs.CV2025
Positional Encoding Field
Yunpeng Bai, Haoxiang Li, Qixing Huang
Diffusion Transformers (DiTs) have emerged as the dominant architecture for visual generation, powering state-of-the-art image and video models. By representing images as patch tok…
cs.CV2025
GeoRemover: Removing Objects and Their Causal Visual Artifacts
Zixin Zhu, Haoxiang Li, Xuelu Feng +3
Towards intelligent image editing, object removal should eliminate both the target object and its causal visual artifacts, such as shadows and reflections. However, existing image…
cs.CV2017★ 17 cited
Learning Dense Facial Correspondences in Unconstrained Images
Ronald Yu, Shunsuke Saito, Haoxiang Li +2
We present a minimalistic but effective neural network that computes dense facial correspondences in highly unconstrained RGB images. Our network learns a per-pixel flow and a matc…