2 papers
cs.LG2026
Neural Variance-aware Dueling Bandits with Deep Representation and Shallow Exploration
Youngmin Oh, Jinje Park, Taejin Paik
We introduce the first variance-aware algorithms for contextual dueling bandits that leverage shallow exploration strategies with neural networks for nonlinear utility approximatio…
cs.LG2024
M3: Mamba-assisted Multi-Circuit Optimization via MBRL with Effective Scheduling
Youngmin Oh, Jinje Park, Seunggeun Kim +3
Recent advancements in reinforcement learning (RL) for analog circuit optimization have demonstrated significant potential for improving sample efficiency and generalization across…