1 paper
Patrick Emami, Nan Qiang, Peter Graf
Supervised fine-tuning (SFT) improves end-to-end classical planning in large language models (LLMs), but do these models also learn to represent and reason about the planning probl…