1 citations · 1 across the 3 of their papers we have counts for
3 papers
cs.CL2024★ 1 cited
DAHL: Domain-specific Automated Hallucination Evaluation of Long-Form Text through a Benchmark Dataset in Biomedicine
Jean Seo, Jongwon Lim, Dongjun Jang +1
We introduce DAHL, a benchmark dataset and automated evaluation system designed to assess hallucination in long-form text generation, specifically within the biomedical domain. Our…
cs.CL2024
Korean Bio-Medical Corpus (KBMC) for Medical Named Entity Recognition
Sungjoo Byun, Jiseung Hong, Sumin Park +5
Named Entity Recognition (NER) plays a pivotal role in medical Natural Language Processing (NLP). Yet, there has not been an open-source medical NER dataset specifically for the Ko…
cs.CL2024
CARBD-Ko: A Contextually Annotated Review Benchmark Dataset for Aspect-Level Sentiment Classification in Korean
Dongjun Jang, Jean Seo, Sungjoo Byun +3
This paper explores the challenges posed by aspect-based sentiment classification (ABSC) within pretrained language models (PLMs), with a particular focus on contextualization and…