feedback-driven policy discovery 1generative recommendation 1intent modeling 1knowledge distillation 1large language models 1online inference 1
From the 1 of 13 linked papers with an AI index.
Showing cs.AIShow all
3 papers · 1 filter
cs.AI2026
Phase-Aware Mixture of Experts for Agentic Reinforcement Learning
Shengtian Yang, Yu Li, Shuo He +4
Reinforcement learning (RL) has equipped LLM agents with a strong ability to solve complex tasks. However, existing RL methods normally use a \emph{single} policy network, causing…
cs.AI2025
GAS: Generative Auto-bidding with Post-training Search
Yewen Li, Shuai Mao, Jingtong Gao +6
Auto-bidding is essential in facilitating online advertising by automatically placing bids on behalf of advertisers. Generative auto-bidding, which generates bids based on an adjus…
cs.AI2024
ACQ: A Deployed Two-Stage Framework for Automated Creative Quota Allocation in Large-Scale Online Advertising
Ruizhi Wang, Kai Liu, Bingjie Li +4
In digital advertising, demand-side platforms (DSPs) allow advertisers to create multiple ad creatives from a single photo for real-time bidding. While increasing the number of cre…