4 citations · 4 across the 1 of their papers we have counts for
1 paper
Yuda Song, Hanlin Zhang, Carson Eisenach +3
Self-improvement is a mechanism in Large Language Model (LLM) pre-training, post-training and test-time inference. We explore a framework where the model verifies its own outputs,…