1 paper · 1 filter
Martin Weyssow, Aton Kamanda, Xin Zhou +1
Evaluating the alignment of large language models (LLMs) with user-defined coding preferences is a challenging endeavour that requires a deep assessment of LLMs' outputs. Existing…