Saskia Redgate, Andrew M. Bean, Adam Mahdi
The growing capabilities of large language models (LLMs) have led to their use as substitutes for human feedback for training and assessing other LLMs. These methods often rely on…