2 papers
cs.CL2026
GRACE: Step-Level Benchmark for Faithful Reasoning over Context
Hoang Pham, Dong Le, Anh Tuan Luu
Many reasoning tasks require models to reason over input context, from document-grounded question answering to rule-based deduction. Chain-of-Thought (CoT) prompting produces trace…
cs.CL2024
Crossing Linguistic Horizons: Finetuning and Comprehensive Evaluation of Vietnamese Large Language Models
Sang T. Truong, Duc Q. Nguyen, Toan Nguyen +4
Recent advancements in large language models (LLMs) have underscored their importance in the evolution of artificial intelligence. However, despite extensive pretraining on multili…