Showing cs.AIShow all
2 papers · 1 filter
cs.AI2025
AgentTTS: Large Language Model Agent for Test-time Compute-optimal Scaling Strategy in Complex Tasks
Fali Wang, Hui Liu, Zhenwei Dai +8
Test-time scaling (TTS) enhances the performance of large language models (LLMs) by allocating additional compute resources during inference. However, existing research primarily i…
cs.AI2024
A Survey of Calibration Process for Black-Box LLMs
Liangru Xie, Hui Liu, Jingying Zeng +7
Large Language Models (LLMs) demonstrate remarkable performance in semantic understanding and generation, yet accurately assessing their output reliability remains a significant ch…