2 papers
cs.AI2026
New Bounds for Zarankiewicz Numbers via Reinforced LLM Evolutionary Search
Jay Bhan, Nicole Nobili, Patrick Langer
The Zarankiewicz number is the maximum number of edges in a bipartite graph such that there is no complete bipartite subgraph. We det…
cs.LG2024
Truncating Trajectories in Monte Carlo Policy Evaluation: an Adaptive Approach
Riccardo Poiani, Nicole Nobili, Alberto Maria Metelli +1
Policy evaluation via Monte Carlo (MC) simulation is at the core of many MC Reinforcement Learning (RL) algorithms (e.g., policy gradient methods). In this context, the designer of…