5 citations · 14 across the 10 of their papers we have counts for
4 papers · 1 filter
Cooperative-Competitive Reinforcement Learning with History-Dependent Rewards
Keyang He, Bikramjit Banerjee, Prashant Doshi
Consider a typical organization whose worker agents seek to collectively cooperate for its general betterment. However, each individual agent simultaneously seeks to act to secure…
Active Deception using Factored Interactive POMDPs to Recognize Cyber Attacker's Intent
Aditya Shinde, Prashant Doshi, Omid Setayeshfar
This paper presents an intelligent and adaptive agent that employs deception to recognize a cyber adversary's intent. Unlike previous approaches to cyber deception, which mainly fo…
Recurrent Sum-Product-Max Networks for Decision Making in Perfectly-Observed Environments
Hari Teja Tatavarti, Prashant Doshi, Layton Hayes
Recent investigations into sum-product-max networks (SPMN) that generalize sum-product networks (SPN) offer a data-driven alternative for decision making, which has predominantly r…
Maximum Entropy Multi-Task Inverse RL
Saurabh Arora, Bikramjit Banerjee, Prashant Doshi
Multi-task IRL allows for the possibility that the expert could be switching between multiple ways of solving the same problem, or interleaving demonstrations of multiple tasks. Th…