15 citations · 16 across the 4 of their papers we have counts for
4 papers · 1 filter
TableVQA-Bench: A Visual Question Answering Benchmark on Multiple Table Domains
Yoonsik Kim, Moonbin Yim, Ka Yeon Song
In this paper, we establish a benchmark for table visual question answering, referred to as the TableVQA-Bench, derived from pre-existing table question-answering (QA) and table st…
SCOB: Universal Text Understanding via Character-wise Supervised Contrastive Learning with Online Text Rendering for Bridging Domain Gap
Daehee Kim, Yoonsik Kim, DongHyun Kim +3
Inspired by the great success of language model (LM)-based pre-training, recent studies in visual document understanding have explored LM-based pre-training methods for modeling te…
Towards Unified Scene Text Spotting based on Sequence Generation
Taeho Kil, Seonghyeon Kim, Sukmin Seo +2
Sequence generation models have recently made significant progress in unifying various vision tasks. Although some auto-regressive models have demonstrated promising results in end…
A New Convolutional Network-in-Network Structure and Its Applications in Skin Detection, Semantic Segmentation, and Artifact Reduction
Yoonsik Kim, Insung Hwang, Nam Ik Cho
The inception network has been shown to provide good performance on image classification problems, but there are not much evidences that it is also effective for the image restorat…