1 paper
John Chen, Sihan Cheng, Can Gurkan +1
Evaluating strategic decision-making in LLM-based agents requires generative, competitive, and longitudinal environments, yet few benchmarks provide all three, and fewer still offe…