2 papers
cs.LG2026
Offline Policy Evaluation as a decision support tool for designing Adaptive Experiments
João Victor Ferreira Alves, Eduardo Rocha Laurentino, Gustavo de Oliveira Kanno +1
We investigate how historical data from fixed randomized experiments (A/B tests) can be used to inform the deployment of adaptive experiments based on contextual bandits. Given dat…
cs.AI2026
OptiAgent: End-to-End Optimization Modeling via Multi-Agent Iterative Refinement
Adriana Laurindo Monteiro, Nayse Fagundes, Gabriel Mattos Langeloh +4
We propose OptiAgent, a multi-agent framework that, given a natural language description of an Operations Research problem, is able to output a solver-ready mathematical formulatio…