activity
20182020
collaborators

7 papers

cs.LG2020

Deep PQR: Solving Inverse Reinforcement Learning using Anchor Actions

Sinong Geng, Houssam Nassif, Carlos A. Manzanares +2

We propose a reward function estimation framework for inverse reinforcement learning with deep energy-based policies. We name our method PQR, as it sequentially estimates the Polic…

cs.IR2019

Seeker: Real-Time Interactive Search

Ari Biswas, Thai T Pham, Michael Vogelsong +2

This paper introduces Seeker, a system that allows users to interactively refine search rankings in real time, through feedback in the form of likes and dislikes. When searching on…

cs.LG2018

An Efficient Bandit Algorithm for Realtime Multivariate Optimization

Daniel N Hill, Houssam Nassif, Yi Liu +2

Optimization is commonly employed to determine the content of web pages, such as to maximize conversions on landing pages or click-through rates on search engine result pages. Ofte…

q-bio.OT2018

An Inductive Logic Programming Approach to Validate Hexose Binding Biochemical Knowledge

Houssam Nassif, Hassan Al-Ali, Sawsan Khuri +2

Hexoses are simple sugars that play a key role in many cellular pathways, and in the regulation of development and disease mechanisms. Current protein-sugar computational models ar…

cs.LG2018

Contextual Multi-Armed Bandits for Causal Marketing

Neela Sawant, Chitti Babu Namballa, Narayanan Sadagopan +1

This work explores the idea of a causal contextual multi-armed bandit approach to automated marketing, where we estimate and optimize the causal (incremental) effects. Focusing on…

cs.MM2018

Diversifying Music Recommendations

Houssam Nassif, Kemal Oral Cansizlar, Mitchell Goodman +1

We compare submodular and Jaccard methods to diversify Amazon Music recommendations. Submodularity significantly improves recommendation quality and user engagement. Unlike the Jac…