1 citations · 1 across the 3 of their papers we have counts for
3 papers
cs.CL2026
FlowLM: Few-Step Language Modeling via Diffusion-to-Flow Adaptation
Runzhe Zhang, Letian Chen, Wenpeng Zhang +2
We present FlowLM, a flow matching language model transformed from pre-trained diffusion language models via efficient fine-tuning. By re-aligning the curved sampling trajectories…
cs.LG2023★ 1 cited
Marketing Budget Allocation with Offline Constrained Deep Reinforcement Learning
Tianchi Cai, Jiyan Jiang, Wenpeng Zhang +7
We study the budget allocation problem in online marketing campaigns that utilize previously collected offline data. We first discuss the long-term effect of optimizing marketing b…
cs.LG2023
Model-free Reinforcement Learning with Stochastic Reward Stabilization for Recommender Systems
Tianchi Cai, Shenliao Bao, Jiyan Jiang +5
Model-free RL-based recommender systems have recently received increasing research attention due to their capability to handle partial feedback and long-term rewards. However, most…