2 papers
cs.AR2025
HEPPO-GAE: Hardware-Efficient Proximal Policy Optimization with Generalized Advantage Estimation
Hazem Taha, Ameer M. S. Abdelhadi
This paper introduces HEPPO-GAE, an FPGA-based accelerator designed to optimize the Generalized Advantage Estimation (GAE) stage in Proximal Policy Optimization (PPO). Unlike previ…
cs.LG2024
Schrödinger's FP: Dynamic Adaptation of Floating-Point Containers for Deep Learning Training
MiloÅ¡ NikoliÄ, Enrique Torres Sanchez, Jiahui Wang +5
The transfer of tensors from/to memory during neural network training dominates time and energy. To improve energy efficiency and performance, research has been exploring ways to u…