41 citations · 75 across the 21 of their papers we have counts for
13 papers · 1 filter
Revisiting Cephalometric Landmark Detection from the view of Human Pose Estimation with Lightweight Super-Resolution Head
Qian Wu, Si Yong Yeo, Yufei Chen +1
Accurate localization of cephalometric landmarks holds great importance in the fields of orthodontics and orthognathics due to its potential for automating key point labeling. In t…
Unsupervised Domain Adaptation via Domain-Adaptive Diffusion
Duo Peng, Qiuhong Ke, Yinjie Lei +1
Unsupervised Domain Adaptation (UDA) is quite challenging due to the large distribution discrepancy between the source domain and the target domain. Inspired by diffusion models wh…
DISGO: Automatic End-to-End Evaluation for Scene Text OCR
Mei-Yuh Hwang, Yangyang Shi, Ankit Ramchandani +6
This paper discusses the challenges of optical character recognition (OCR) on natural scenes, which is harder than OCR on documents due to the wild content and various image backgr…
Diffusion-based Image Translation with Label Guidance for Domain Adaptive Semantic Segmentation
Duo Peng, Ping Hu, Qiuhong Ke +1
Translating images from a source domain to a target domain for learning target models is one of the most common strategies in domain adaptive semantic segmentation (DASS). However,…
AI-Generated Content (AIGC) for Various Data Modalities: A Survey
Lin Geng Foo, Hossein Rahmani, Jun Liu
AI-generated content (AIGC) methods aim to produce text, images, videos, 3D assets, and other media using AI algorithms. Due to its wide range of applications and the potential of…
Distribution-Aligned Diffusion for Human Mesh Recovery
Lin Geng Foo, Jia Gong, Hossein Rahmani +1
Recovering a 3D human mesh from a single RGB image is a challenging task due to depth ambiguity and self-occlusion, resulting in a high degree of uncertainty. Meanwhile, diffusion…