1 paper · 1 filter
Sina Alemohammad, Li Chen, Richard G. Baraniuk +1
Can a language model improve from plain text sampled from itself, with no prompts, no teacher, no verifier, and no reward model? Yes, but only when the synthetic corpus is compatib…