3 citations · 3 across the 2 of their papers we have counts for
2 papers
cs.CL2024
HammerBench: Fine-Grained Function-Calling Evaluation in Real Mobile Device Scenarios
Jun Wang, Jiamu Zhou, Muning Wen +7
Evaluating the performance of LLMs in multi-turn human-agent interactions presents significant challenges, particularly due to the complexity and variability of user behavior. In t…
cs.AI2023★ 3 cited
Order Matters: Agent-by-agent Policy Optimization
Xihuai Wang, Zheng Tian, Ziyu Wan +3
While multi-agent trust region algorithms have achieved great success empirically in solving coordination tasks, most of them, however, suffer from a non-stationarity problem since…