1 paper
Ruizhe Chen, Wenhao Chai, Zhifei Yang +5
Inference-time alignment provides an efficient alternative for aligning LLMs with humans. However, these approaches still face challenges, such as limited scalability due to policy…