2 citations · 6 across the 8 of their papers we have counts for
Showing cs.SEShow all
2 papers · 1 filter
cs.SE2025
Live API-Bench: 2500+ Live APIs for Testing Multi-Step Tool Calling
Benjamin Elder, Anupama Murthi, Jungkoo Kang +4
Large language models (LLMs) increasingly rely on external tools and APIs to execute complex tasks specified in natural language. Evaluating such tool calling capabilities in reali…
cs.SE2024★ 2 cited
On the Standardization of Behavioral Use Clauses and Their Adoption for Responsible Licensing of AI
Daniel McDuff, Tim Korjakow, Scott Cambo +9
Growing concerns over negligent or malicious uses of AI have increased the appetite for tools that help manage the risks of the technology. In 2018, licenses with behaviorial-use c…