Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
CoBRA: Learning Tool-Use Boundaries via Counterfactual Margins
Wenhao Zou, Xianglong Liu, Wendong Bi +3
As large language models increasingly act through external tools, deciding when to call a tool has become a central problem alongside deciding how to use it. Unnecessary tool calls…
cs.AI2025
ContextPRM: Leveraging Contextual Coherence for multi-domain Test-Time Scaling
Haotian Zhang, Liu Liu, Baosheng Yu +5
Process reward models (PRMs) have demonstrated significant efficacy in enhancing the mathematical reasoning capabilities of large language models (LLMs) by leveraging test-time sca…