1 paper
Xuan He, Da Yin, Nanyun Peng
How can "weak teacher models" such as average human annotators or existing AI systems, effectively supervise LLMs to improve performance on hard reasoning tasks, especially those t…