3 papers
cs.LG2026
Rethinking Privileged Information in On-Policy Self-Distillation
Samyak Shrestha, Alexander Tessier
On-policy self-distillation (OPSD) trains a student on its own responses using token-level supervision from the same model conditioned on privileged reference information. We inves…
cs.AI2026
DeepSciVerify: Verifying Scientific Claim--Citation Alignment via LLM-Driven Evidence Escalation
Shaghayegh Sadeghi, Khashayar Khajavi, Rise Adhikari +1
Misalignment between claims and their cited evidence is a common failure mode in reports generated by large language models, limiting their reliability in scientific and other high…
cs.DL2026
CiteCheck: Retrieval-Grounded Detection of LLM Citation Hallucinations in Scientific Text
Khashayar Khajavi, Shaghayegh Sadeghi, Rise Adhikari +1
Large language models (LLMs) are increasingly used to generate scientific reports, but they can produce references that appear plausible while containing corrupted metadata or poin…