1 paper
Jiashu He, Mingyu Derek Ma, Jinxuan Fan +3
Existing approaches based on context prompting or reinforcement learning (RL) to improve the reasoning capacities of large language models (LLMs) depend on the LLMs' internal knowl…