3 papers
cs.CL2026
Bridging Reasoning Trajectories in On-Policy Distillation via Near-Future Guidance
Yuxuan Jiang, Francis Ferraro
On-Policy Distillation (OPD) improves large language model reasoning by training a student model on trajectories sampled from its own policy under teacher supervision. Although OPD…
cs.CL2026
Experiments or Outcomes? Probing Scientific Feasibility in Large Language Models
Seyedali Mohammadi, Manas Gaur, Francis Ferraro
Scientific feasibility assessment asks whether a claim is consistent with established knowledge and whether experimental evidence could support or refute it. We frame feasibility a…
cs.CL2025
FRIDA to the Rescue! Analyzing Synthetic Data Effectiveness in Object-Based Common Sense Reasoning for Disaster Response
Mollie Shichman, Claire Bonial, Austin Blodgett +3
During Human Robot Interactions in disaster relief scenarios, Large Language Models (LLMs) have the potential for substantial physical reasoning to assist in mission objectives. Ho…