2 papers
cs.AI2026
TriQua: Reconciling Granularity and Context in Factuality Evaluation
Jin Liu, Steffen Thoma, Achim Rettinger
The "decompose-then-verify" paradigm for LLM factuality evaluation faces a fundamental trade-off: atomic facts, i.e., one sentence conveying one unit of information, often omit ess…
cs.CL2024
FZI-WIM at SemEval-2024 Task 2: Self-Consistent CoT for Complex NLI in Biomedical Domain
Jin Liu, Steffen Thoma
This paper describes the inference system of FZI-WIM at the SemEval-2024 Task 2: Safe Biomedical Natural Language Inference for Clinical Trials. Our system utilizes the chain of th…