2 papers
cs.CL2026
End-to-End Chatbot Evaluation with Adaptive Reasoning and Uncertainty Filtering
Nhi Dang, Tung Le, Huy Tien Nguyen
Large language models (LLMs) combined with retrieval augmented generation have enabled the deployment of domain-specific chatbots, but these systems remain prone to generating unsu…
cs.CL2026
From Prompting to Preference Optimization: A Comparative Study of LLM-based Automated Essay Scoring
Minh Hoang Nguyen, Vu Hoang Pham, Xuan Thanh Huynh +5
Large language models (LLMs) have recently reshaped Automated Essay Scoring (AES), yet prior studies typically examine individual techniques in isolation, limiting understanding of…