1 paper · 1 filter
Bolian Li, Yanran Wu, Xinyu Luo +1
Aligning large language models (LLMs) with human preferences has become a critical step in their development. Recent research has increasingly focused on test-time alignment, where…