29 citations · 57 across the 7 of their papers we have counts for
Showing 2018Show all
2 papers · 1 filter
stat.ML2018
Simple Regret Minimization for Contextual Bandits
Aniket Anand Deshmukh, Srinagesh Sharma, James W. Cutler +2
There are two variants of the classical multi-armed bandit (MAB) problem that have received considerable attention from machine learning researchers in recent years: contextual ban…
cs.LG2018
Domain2Vec: Deep Domain Generalization
Aniket Anand Deshmukh, Ankit Bansal, Akash Rastogi
We address the problem of domain generalization where a decision function is learned from the data of several related domains, and the goal is to apply it on an unseen domain succe…