Showing cs.AIShow all
2 papers · 1 filter
cs.AI2025
RATIONALYST: Mining Implicit Rationales for Process Supervision of Reasoning
Dongwei Jiang, Guoxuan Wang, Yining Lu +5
The reasoning steps generated by LLMs might be incomplete, as they mimic logical leaps common in everyday communication found in their pre-training data: underlying rationales are…
cs.AI2024
SELF-[IN]CORRECT: LLMs Struggle with Discriminating Self-Generated Responses
Dongwei Jiang, Jingyu Zhang, Orion Weller +3
Can LLMs consistently improve their previous outputs for better results? For this to be true, LLMs would need to be better at discriminating among previously-generated alternatives…