1 paper · 1 filter
Chengqi Zheng, Keya Hu, Shuzhi Liu +3
LLM prompting is widely used for naturally stated tasks, yet it is unreliable it may succeed on a few test cases but fail at deployment time. We study performance prediction: given…