10 citations · 10 across the 4 of their papers we have counts for
5 papers
R4-CGQA: Retrieval-based Vision Language Models for Computer Graphics Image Quality Assessment
Zhuangzi Li, Jian Jin, Shilv Cai +1
Immersive Computer Graphics (CGs) rendering has become ubiquitous in modern daily life. However, comprehensively evaluating CG quality remains challenging for two reasons: First, e…
A Large Cross-Modal Video Retrieval Dataset with Reading Comprehension
Weijia Wu, Yuzhong Zhao, Zhuang Li +4
Most existing cross-modal language-to-video retrieval (VR) research focuses on single-modal input from video, i.e., visual representation, while the text is omnipresent in human en…
FlowText: Synthesizing Realistic Scene Text Video with Optical Flow Estimation
Yuzhong Zhao, Weijia Wu, Zhuang Li +2
Current video text spotting methods can achieve preferable performance, powered with sufficient labeled training data. However, labeling data manually is time-consuming and labor-i…
ICDAR 2023 Video Text Reading Competition for Dense and Small Text
Weijia Wu, Yuzhong Zhao, Zhuang Li +5
Recently, video text detection, tracking, and recognition in natural scenes are becoming very popular in the computer vision community. However, most existing algorithms and benchm…
Image Super-Resolution Using Attention Based DenseNet with Residual Deconvolution
Zhuangzi Li
Image super-resolution is a challenging task and has attracted increasing attention in research and industrial communities. In this paper, we propose a novel end-to-end Attention-b…