3 papers
cs.LG2026
K-FinHallu: A Hallucination Detection Benchmark for Multi-Turn RAG in Korean Finance
Eunbyeol Cho, Yunseung Lee, Mirae Kim +3
Large Language Models (LLMs) have advanced financial automation through Retrieval-Augmented Generation (RAG), yet hallucinations remain a critical barrier to deployment in high-sta…
cs.LG2026
Generating Multi-Table Time Series EHR from Latent Space with Minimal Preprocessing
Eunbyeol Cho, Jiyoun Kim, Minjae Lee +2
Electronic Health Records (EHR) are time-series relational databases that record patient interactions and medical events over time, serving as a critical resource for healthcare re…
cs.CL2025
DialSim: A Dialogue Simulator for Evaluating Long-Term Multi-Party Dialogue Understanding of Conversational Agents
Jiho Kim, Woosog Chay, Hyeonji Hwang +6
Recent advancements in Large Language Models (LLMs) have significantly enhanced conversational agents, making them applicable to various fields (e.g., education, entertainment). De…