17 citations · 27 across the 5 of their papers we have counts for
5 papers
BRAC+: Improved Behavior Regularized Actor Critic for Offline Reinforcement Learning
Chi Zhang, Sanmukh Rao Kuppannagari, Viktor K Prasanna
Online interactions with the environment to collect data samples for training a Reinforcement Learning (RL) agent is not always feasible due to economic and safety concerns. The go…
A High Throughput Parallel Hash Table on FPGA using XOR-based Memory
Ruizhi Zhang, Sasindu Wijeratne, Yang Yang +2
Hash table is a fundamental data structure for quick search and retrieval of data. It is a key component in complex graph analytics and AI/ML applications. State-of-the-art paralle…
DYNAMAP: Dynamic Algorithm Mapping Framework for Low Latency CNN Inference
Yuan Meng, Sanmukh Kuppannagari, Rajgopal Kannan +1
Most of the existing work on FPGA acceleration of Convolutional Neural Network (CNN) focus on employing a single strategy (algorithm, dataflow, etc.) across all the layers. Such an…
Maximum Entropy Model Rollouts: Fast Model Based Policy Optimization without Compounding Errors
Chi Zhang, Sanmukh Rao Kuppannagari, Viktor K Prasanna
Model usage is the central challenge of model-based reinforcement learning. Although dynamics model based on deep neural networks provide good generalization for single step predic…
Building HVAC Scheduling Using Reinforcement Learning via Neural Network Based Model Approximation
Chi Zhang, Sanmukh R. Kuppannagari, Rajgopal Kannan +1
Buildings sector is one of the major consumers of energy in the United States. The buildings HVAC (Heating, Ventilation, and Air Conditioning) systems, whose functionality is to ma…