2 papers
cs.CL2026
Knowledge Graph-Assisted LLM Post-Training for Enhanced Legal Reasoning
Dezhao Song, Guglielmo Bonifazi, Frank Schilder +1
LLM post-training has primarily relied on large text corpora and human feedback, without capturing the structure of domain knowledge. This has caused models to struggle dealing wit…
cs.CL2025
Beyond Pointwise Scores: Decomposed Criteria-Based Evaluation of LLM Responses
Fangyi Yu, Nabeel Seedat, Dasha Herrmannova +2
Evaluating long-form answers in high-stakes domains such as law or medicine remains a fundamental challenge. Standard metrics like BLEU and ROUGE fail to capture semantic correctne…