4 citations · 4 across the 2 of their papers we have counts for
5 papers
FreeEnhance: Tuning-Free Image Enhancement via Content-Consistent Noising-and-Denoising Process
Yang Luo, Yiheng Zhang, Zhaofan Qiu +4
The emergence of text-to-image generation models has led to the recognition that image enhancement, performed as post-processing, would significantly improve the visual quality of…
3DStyle-Diffusion: Pursuing Fine-grained Text-driven 3D Stylization with 2D Diffusion Models
Haibo Yang, Yang Chen, Yingwei Pan +3
3D content creation via text-driven stylization has played a fundamental challenge to multimedia and graphics community. Recent advances of cross-modal foundation models (e.g., CLI…
TPS++: Attention-Enhanced Thin-Plate Spline for Scene Text Recognition
Tianlun Zheng, Zhineng Chen, Jinfeng Bai +2
Text irregularities pose significant challenges to scene text recognizers. Thin-Plate Spline (TPS)-based rectification is widely regarded as an effective means to deal with them. C…
Web Video Categorization based on Wikipedia Categories and Content-Duplicated Open Resources
Zhineng Chen, Juan Cao, Yicheng Song +2
This paper presents a novel approach for web video categorization by leveraging Wikipedia categories (WikiCs) and open resources describing the same content as the video, i.e., con…
Context-Oriented Web Video Tag Recommendation
Zhineng Chen, Juan Cao, Yicheng Song +3
Tag recommendation is a common way to enrich the textual annotation of multimedia contents. However, state-of-the-art recommendation methods are built upon the pair-wised tag relev…