3 papers
cs.LG2026
FERPO: Forward Entropy-Regularized Policy Optimization
Sebastian Sanokowski, Alireza Sarmadi, Majid Khadiv
Several state-of-the-art methods for online reinforcement learning in continuous control improve policies using action gradients of a learned critic. However, critics are typically…
cs.CR2026
Long-Term and Short-Term Transistor Aging in Deep Neural Networks: Impact and Mitigation
Alireza Sarmadi, Virinchi Roy Surabhi, Prashanth Krishnamurthy +3
Deep neural networks (DNNs) are used in a variety of real-world applications including, for example, image classification and speech recognition. The inference accuracy of DNN impl…
eess.SY2023
High-Dimensional Controller Tuning through Latent Representations
Alireza Sarmadi, Prashanth Krishnamurthy, Farshad Khorrami
In this paper, we propose a method to automatically and efficiently tune high-dimensional vectors of controller parameters. The proposed method first learns a mapping from the high…