1 paper · 1 filter
Wenhao Yuan, Chenchen Lin, Jian Chen +3
In black-box large language model (LLM) services, response reliability is often only partially observable at decision time, while stronger inference pathways incur substantial comp…