3 papers
cs.LG2026
COOPA: A Modular LLM Agent Architecture for Operations Research Problems
Chuanhao Li, Xiaoan Xu, Dirk Bergemann +3
Operations Research (OR) provides a rigorous framework for high-stakes decision-making, but effective OR modeling requires substantial domain knowledge, mathematical abstraction, a…
cs.SE2026
LongCLI-Bench: A Preliminary Benchmark and Study for Long-horizon Agentic Programming in Command-Line Interfaces
Yukang Feng, Jianwen Sun, Zelai Yang +16
Recent advances in AI-assisted programming have empowered agents to execute complex workflows via command-line interfaces, however, existing benchmarks are limited by short task ho…
cs.AI2025
Build Your Personalized Research Group: A Multiagent Framework for Continual and Interactive Science Automation
Ed Li, Junyu Ren, Xintian Pan +4
The automation of scientific discovery represents a critical milestone in Artificial Intelligence (AI) research. However, existing agentic systems for science suffer from two funda…