2 papers
cs.CL2025
Integrated Framework for LLM Evaluation with Answer Generation
Sujeong Lee, Hayoung Lee, Seongsoo Heo +1
Reliable evaluation of large language models is essential to ensure their applicability in practical scenarios. Traditional benchmark-based evaluation methods often rely on fixed r…
cs.CL2025
HuDEx: Integrating Hallucination Detection and Explainability for Enhancing the Reliability of LLM responses
Sujeong Lee, Hayoung Lee, Seongsoo Heo +1
Recent advances in large language models (LLMs) have shown promising improvements, often surpassing existing methods across a wide range of downstream tasks in natural language pro…