1 paper
Guy Tennenholtz, Yinlam Chow, Chih-Wei Hsu +3
We propose a novel approach for training large language models (LLMs) to adhere to objectives defined within a latent embedding space. Our method leverages reinforcement learning (…