1 paper · 1 filter
Vaibhav Prakash, Jayasri Dontabhaktuni
Language models fine-tuned where the correct completion must outrank a near-synonym competitor often fail silently. The cross-entropy loss falls monotonically while the correct tok…