2 papers
cs.AI2026
Not All Errors Are Equal: Consequence-Aware Reasoning Compute Allocation
Jingbo Wen, Liang He, Ziqi He
Modern reasoning models can allocate different amounts of test-time computation, such as thinking tokens, model calls, or compute budget, to different tasks. Existing methods gener…
cs.SE2024
Evaluation and Improvement of Fault Detection for Large Language Models
Qiang Hu, Jin Wen, Maxime Cordy +4
Large language models (LLMs) have recently achieved significant success across various application domains, garnering substantial attention from different communities. Unfortunatel…