Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
FrontierOR: Benchmarking LLMs' Capacity for Efficient Algorithm Design in Large-Scale Optimization
Minwei Kong, Chonghe Jiang, Ao Qu +24
Large language models (LLMs) are increasingly used for optimization modeling and solver-code generation, yet practical operations research and optimization problems often require a…
cs.AI2025
HugAgent: Benchmarking LLMs for Simulation of Individualized Human Reasoning
Chance Jiajie Li, Zhenze Mo, Yuhan Tang +11
Simulating human reasoning in open-ended tasks has long been a central aspiration in AI and cognitive science. While large language models now approximate human responses at scale,…