7 papers
Deep PQR: Solving Inverse Reinforcement Learning using Anchor Actions
Sinong Geng, Houssam Nassif, Carlos A. Manzanares +2
We propose a reward function estimation framework for inverse reinforcement learning with deep energy-based policies. We name our method PQR, as it sequentially estimates the Polic…
Seeker: Real-Time Interactive Search
Ari Biswas, Thai T Pham, Michael Vogelsong +2
This paper introduces Seeker, a system that allows users to interactively refine search rankings in real time, through feedback in the form of likes and dislikes. When searching on…
An Efficient Bandit Algorithm for Realtime Multivariate Optimization
Daniel N Hill, Houssam Nassif, Yi Liu +2
Optimization is commonly employed to determine the content of web pages, such as to maximize conversions on landing pages or click-through rates on search engine result pages. Ofte…
An Inductive Logic Programming Approach to Validate Hexose Binding Biochemical Knowledge
Houssam Nassif, Hassan Al-Ali, Sawsan Khuri +2
Hexoses are simple sugars that play a key role in many cellular pathways, and in the regulation of development and disease mechanisms. Current protein-sugar computational models ar…
Contextual Multi-Armed Bandits for Causal Marketing
Neela Sawant, Chitti Babu Namballa, Narayanan Sadagopan +1
This work explores the idea of a causal contextual multi-armed bandit approach to automated marketing, where we estimate and optimize the causal (incremental) effects. Focusing on…
Diversifying Music Recommendations
Houssam Nassif, Kemal Oral Cansizlar, Mitchell Goodman +1
We compare submodular and Jaccard methods to diversify Amazon Music recommendations. Submodularity significantly improves recommendation quality and user engagement. Unlike the Jac…