2 papers
cs.LG2025
Few-Shot Test-Time Optimization Without Retraining for Semiconductor Recipe Generation and Beyond
Shangding Gu, Donghao Ying, Ming Jin +4
We introduce Model Feedback Learning (MFL), a novel test-time optimization framework for optimizing inputs to pre-trained AI models or deployed hardware systems without requiring a…
cs.LG2025
Don't Trade Off Safety: Diffusion Regularization for Constrained Offline RL
Junyu Guo, Zhi Zheng, Donghao Ying +4
Constrained reinforcement learning (RL) seeks high-performance policies under safety constraints. We focus on an offline setting where the agent has only a fixed dataset -- common…