4 papers
ClinOCR-Bench: A Comprehensive Clinical Scanned Document Dataset for Optical Character Recognition Model Evaluation
Enshuo Hsu, Jin Zhou, Kirk Roberts
Extracting textual information from scanned medical documents, such as external laboratory reports and manually filled forms, has been a major challenge in modern electronic health…
Synthesized Annotation Guidelines are Knowledge-Lite Boosters for Clinical Information Extraction
Enshuo Hsu, Martin Ugbala, Krishna Kumar Kookal +4
Generative information extraction using large language models, particularly through few-shot learning, has become a popular method. Recent studies indicate that providing a detaile…
LLM-IE: A Python Package for Generative Information Extraction with Large Language Models
Enshuo Hsu, Kirk Roberts
Objectives: Despite the recent adoption of large language models (LLMs) for biomedical information extraction, challenges in prompt engineering and algorithms persist, with no dedi…
Leveraging Large Language Models for Knowledge-free Weak Supervision in Clinical Natural Language Processing
Enshuo Hsu, Kirk Roberts
The performance of deep learning-based natural language processing systems is based on large amounts of labeled training data which, in the clinical domain, are not easily availabl…