1 paper
Sai Wang, Senthilnathan Subramanian, Mudit Sahni +6
Large-language-model (LLM) agents exhibit complex, context-sensitive behaviour that quickly renders static benchmarks and ad-hoc manual testing obsolete. We present Neo, a configur…