1 paper · 1 filter
Derin Cayir, Renjie Tao, Rashi Rungta +6
Large Language Models (LLMs) have demonstrated remarkable progress through preference-based fine-tuning, which critically depends on the quality of the underlying training data. Wh…