From the 1 of 14 papers with an AI index.
5 citations
- Beijing Institute of TechnologyCN2 papers
- University of Science and Technology of ChinaCN2 papers
- Australian National UniversityAU1 paper
- Beijing Jiaotong UniversityCN1 paper
- Chinese Academy of SciencesCN1 paper
- Cloud Computing CenterCN1 paper
- Hangzhou Normal UniversityCN1 paper
- He Eye HospitalCN1 paper
- Hong Kong Polytechnic UniversityHK1 paper
- Institute of MicroelectronicsCN1 paper
- Jingdong (China)CN1 paper
- Macquarie UniversityAU1 paper
4 papers · 1 filter
Visual Information Extraction from Documents via Classification-Guided Large Vision-Language Models
Huafu Li, Guo Chen, Jia Xia +5
Visual information extraction (VIE) from visually rich documents remains challenging due to high layout variability and real-world impairments. Existing methods typically rely on s…
An automated method of identifying incorrectly labelled images based on the sequences of loss functions of deep learning networks
Zhipeng Zhang, Wenhui Shou, Wengting Ma +5
Deep learning is widely applied in medical image analysis, but up to 10% of manually labelled images may be incorrect, degrading model performance. This paper proposes an automated…
EchoSR: Efficient Context Harnessing for Lightweight Image Super-Resolution
Hanli Zhao, Binhao Wang, Shihao Zhao +3
Image super-resolution (SR) aims to reconstruct high-quality, high-resolution (HR) images from low-resolution (LR) inputs and plays a critical role in various downstream applicatio…
DepthMaster: Taming Diffusion Models for Monocular Depth Estimation
Ziyang Song, Zerong Wang, Bo Li +5
Monocular depth estimation within the diffusion-denoising paradigm demonstrates impressive generalization ability but suffers from low inference speed. Recent methods adopt a singl…