1 paper
Jonathan Cook, Tim Rocktäschel, Jakob Foerster +2
Given the widespread adoption and usage of Large Language Models (LLMs), it is crucial to have flexible and interpretable evaluations of their instruction-following ability. Prefer…