activity
20212023
most citedCharFormer: A Glyph Fusion based Attentive Framework for High-precision Character Image Denoising

31 citations · 74 across the 13 of their papers we have counts for

collaborators
Showing cs.CVShow all

22 papers · 1 filter

cs.CV20241 cited

HandDiff: 3D Hand Pose Estimation with Diffusion on Image-Point Cloud

Wencan Cheng, Hao Tang, Luc Van Gool +1

Extracting keypoint locations from input hand frames, known as 3D hand pose estimation, is a critical task in various human-computer interaction applications. Essentially, the 3D h…

cs.CV2024

Towards Online Real-Time Memory-based Video Inpainting Transformers

Guillaume Thiry, Hao Tang, Radu Timofte +1

Video inpainting tasks have seen significant improvements in recent years with the rise of deep neural networks and, in particular, vision transformers. Although these models show…

cs.CV2024

ICON: Incremental CONfidence for Joint Pose and Radiance Field Optimization

Weiyao Wang, Pierre Gleize, Hao Tang +3

Neural Radiance Fields (NeRF) exhibit remarkable performance for Novel View Synthesis (NVS) given a set of 2D images. However, NeRF training requires accurate camera pose for each…

cs.CV2024

Graph Transformer GANs with Graph Masked Modeling for Architectural Layout Generation

Hao Tang, Ling Shao, Nicu Sebe +1

We present a novel graph Transformer generative adversarial network (GTGAN) to learn effective graph node relations in an end-to-end fashion for challenging graph-constrained archi…

cs.CV202317 cited

Towards High-quality HDR Deghosting with Conditional Diffusion Models

Qingsen Yan, Tao Hu, Yuan Sun +5

High Dynamic Range (HDR) images can be recovered from several Low Dynamic Range (LDR) images by existing Deep Neural Networks (DNNs) techniques. Despite the remarkable progress, DN…

cs.CV20232 cited

Efficient-3DiM: Learning a Generalizable Single-image Novel-view Synthesizer in One Day

Yifan Jiang, Hao Tang, Jen-Hao Rick Chang +3

The task of novel view synthesis aims to generate unseen perspectives of an object or scene from a limited set of input images. Nevertheless, synthesizing novel views from a single…