504 citations · 1.3k across the 47 of their papers we have counts for
Showing 2025 · cs.LGShow all
2 papers · 2 filters
cs.LG2025
A result relating convex n-widths to covering numbers with some applications to neural networks
Jonathan Baxter, Peter Bartlett
In general, approximating classes of functions defined over high-dimensional input spaces by linear combinations of a fixed set of basis functions or ``features'' is known to be ha…
cs.LG2025★ 88 cited
Reinforcement Learning in POMDP's via Direct Gradient Ascent
Jonathan Baxter, Peter L. Bartlett
This paper discusses theoretical and experimental aspects of gradient-based approaches to the direct optimization of policy performance in controlled POMDPs. We introduce GPOMDP, a…