2 papers
cs.AI2026
RIFT: Reordered Instruction Following Testbed To Evaluate Instruction Following in Singular Multistep Prompt Structures
Andrew Jaffe, Noah Reicin, Jinho D. Choi
Large Language Models (LLMs) are increasingly relied upon for complex workflows, yet their ability to maintain flow of instructions remains underexplored. Existing benchmarks confl…
cs.CL2025
TRUST: An LLM-Based Dialogue System for Trauma Understanding and Structured Assessments
Sichang Tu, Abigail Powers, Stephen Doogan +1
Objectives: While Large Language Models (LLMs) have been widely used to assist clinicians and support patients, no existing work has explored dialogue systems for standard diagnost…