858 citations · 1.1k across the 4 of their papers we have counts for
5 papers · 1 filter
Index-Preserving Lightweight Token Pruning for Efficient Document Understanding in Vision-Language Models
Jaemin Son, Sujin Choi, Inyong Yun
Recent progress in vision-language models (VLMs) has led to impressive results in document understanding tasks, but their high computational demands remain a challenge. To mitigate…
Character decomposition to resolve class imbalance problem in Hangul OCR
Geonuk Kim, Jaemin Son, Kanghyu Lee +1
We present a novel approach to OCR(Optical Character Recognition) of Korean character, Hangul. As a phonogram, Hangul can represent 11,172 different characters with only 52 graphem…
REFUGE Challenge: A Unified Framework for Evaluating Automated Methods for Glaucoma Assessment from Fundus Photographs
José Ignacio Orlando, Huazhu Fu, João Barbossa Breda +28
Glaucoma is one of the leading causes of irreversible but preventable blindness in working age populations. Color fundus photography (CFP) is the most cost-effective imaging modali…
Classification of Findings with Localized Lesions in Fundoscopic Images using a Regionally Guided CNN
Jaemin Son, Woong Bae, Sangkeun Kim +2
Fundoscopic images are often investigated by ophthalmologists to spot abnormal lesions to make diagnoses. Recent successes of convolutional neural networks are confined to diagnose…
Retinal Vessel Segmentation in Fundoscopic Images with Generative Adversarial Networks
Jaemin Son, Sang Jun Park, Kyu-Hwan Jung
Retinal vessel segmentation is an indispensable step for automatic detection of retinal diseases with fundoscopic images. Though many approaches have been proposed, existing method…