4 papers
Bridging Reasoning Trajectories in On-Policy Distillation via Near-Future Guidance
Yuxuan Jiang, Francis Ferraro
On-Policy Distillation (OPD) improves large language model reasoning by training a student model on trajectories sampled from its own policy under teacher supervision. Although OPD…
Findings of the MAGMaR 2026 Shared Task
Alexander Martin, Dengjia Zhang, Joel Brogan +7
This overview paper presents the results of the shared task for the second workshop on Multimodal Augmented Generation via Multimodal Retrieval (MAGMaR). In this shared task partic…
Experiments or Outcomes? Probing Scientific Feasibility in Large Language Models
Seyedali Mohammadi, Manas Gaur, Francis Ferraro
Scientific feasibility assessment asks whether a claim is consistent with established knowledge and whether experimental evidence could support or refute it. We frame feasibility a…
FRIDA to the Rescue! Analyzing Synthetic Data Effectiveness in Object-Based Common Sense Reasoning for Disaster Response
Mollie Shichman, Claire Bonial, Austin Blodgett +3
During Human Robot Interactions in disaster relief scenarios, Large Language Models (LLMs) have the potential for substantial physical reasoning to assist in mission objectives. Ho…