3 papers
cs.CL2026
Not All Claims Are Equally Risky: FACTOR for Adaptive Verification in Factual Long-Form Generation
Areeba Hassan, Arooj Kausar, Syeda Kisaa Fatima +2
Large Language Models (LLMs) generate fluent long-form text, however, often add unsupported factual claims. Existing verification techniques improve factuality by grounding generat…
cs.AI2026
PRIME: Evaluating Prompt Resolution Under Incompatible Instructions in LLMs
Tehreem Javed, Shumaim Fatimah, Masooma Bakhtiari +2
Large language models (LLMs) often encounter conflicting prompts, although current instruction following benchmarks assess those meta-instructions in isolation, limiting the insigh…
cs.CL2025
Subjective Question Generation and Answer Evaluation using NLP
G. M. Refatul Islam, Safwan Shaheer, Yaseen Nur +1
Natural Language Processing (NLP) is one of the most revolutionary technologies today. It uses artificial intelligence to understand human text and spoken words. It is used for tex…