From the 1 of 5 linked papers with an AI index.
5 papers
Building a User Foundation Model for the Open Web
Solal Vernier, Ivan Can Arisoy, Merwan Barlier +1
The paper introduces a self‑supervised transformer model trained on fragmented web browsing histories to create user representations that improve click prediction and bidding perfo…
PromptPack: Scaling LLM Annotation Agents for Online Recommendation
Sebastian Koralewski, Merwan Barlier, Yulia Stolin +1
Online recommendation platforms increasingly use Large Language Models (LLMs) to extract structured features from ad creatives. While deploying a single-call LLM annotation agent y…
Adaptive Sample Sharing for Multi Agent Linear Bandits
Hamza Cherkaoui, Merwan Barlier, Igor Colin
The multi-agent linear bandit setting is a well-known setting for which designing efficient collaboration between agents remains challenging. This paper studies the impact of data…
Differentially Private Policy Gradient
Alexandre Rio, Merwan Barlier, Igor Colin
Motivated by the increasing deployment of reinforcement learning in the real world, involving a large consumption of personal data, we introduce a differentially private (DP) polic…
Differentially Private Deep Model-Based Reinforcement Learning
Alexandre Rio, Merwan Barlier, Igor Colin +1
We address private deep offline reinforcement learning (RL), where the goal is to train a policy on standard control tasks that is differentially private (DP) with respect to indiv…