107 citations · 165 across the 23 of their papers we have counts for
4 papers · 1 filter
Interpretable Multi-Objective Reinforcement Learning through Policy Orchestration
Ritesh Noothigattu, Djallel Bouneffouf, Nicholas Mattei +6
Autonomous cyber-physical agents and systems play an increasingly large role in our lives. To ensure that agents behave in ways aligned with the values of the societies in which th…
Incorporating Behavioral Constraints in Online AI Systems
Avinash Balakrishnan, Djallel Bouneffouf, Nicholas Mattei +1
AI systems that learn through reward feedback about the actions they take are increasingly deployed in domains that have significant impact on our daily life. However, in many case…
Beyond Backprop: Online Alternating Minimization with Auxiliary Variables
Anna Choromanska, Benjamin Cowen, Sadhana Kumaravel +8
Despite significant recent advances in deep neural networks, training them remains a challenge due to the highly non-convex nature of the objective function. State-of-the-art metho…
Contextual Bandit with Adaptive Feature Extraction
Baihan Lin, Djallel Bouneffouf, Guillermo Cecchi +1
We consider an online decision making setting known as contextual bandit problem, and propose an approach for improving contextual bandit performance by using an adaptive feature e…