2 papers
cs.CL2025
VM14K: First Vietnamese Medical Benchmark
Thong Nguyen, Duc Nguyen, Minh Dang +6
Medical benchmarks are indispensable for evaluating the capabilities of language models in healthcare for non-English-speaking communities,therefore help ensuring the quality of re…
cs.CL2023
Reducing LLM Hallucinations using Epistemic Neural Networks
Shreyas Verma, Kien Tran, Yusuf Ali +1
Reducing and detecting hallucinations in large language models is an open research problem. In this project, we attempt to leverage recent advances in the field of uncertainty esti…