12 papers
DecoEvo: Score-Decoupled Co-Evolution of Solver and Rubric-Generator Skills in Text Space
Jiangwang Chen, Zixin Song, Junlin Liu +10
The paper introduces DecoEvo, a method that co-evolves a solver and a rubric-generator for large language models in text space using decoupled objectives, allowing the solver to im…
MuChator: Enabling Active Music Discovery via Conversational Music LLMs in Douyin Music
Jiahao Liang, Linzhi Huang, Xuannan Liu +6
Douyin Music, a large-scale platform with millions of daily users, adopts an immersive, feed-based discovery paradigm, where users passively explore music through continuous recomm…
OnePred: Next-Query Prediction via Recursive Intent Memory in Multi-Turn Conversations
Jiangwang Chen, Bowen Zhang, Zixin Song +4
Although large language model (LLM) conversational systems process millions of multi-turn dialogues daily, they remain fundamentally reactive: they respond only after the user type…
Co-ReAct: Rubrics as Step-Level Collaborators for ReAct Agents
Jiazheng Kang, Bowen Zhang, Zixin Song +4
ReAct-style agents for search-intensive, multi-step reasoning tasks rely largely on their own internal judgment to decide what evidence to seek, which reasoning or action step to t…
Make It Long, Keep It Fast: End-to-End 10K Long User Behavior Sequence Modeling for Billion-Scale Douyin Recommendation
Lin Guan, Jia-Qi Yang, Zhishan Zhao +12
Short-video recommenders such as Douyin must exploit extremely long user behavior histories without breaking latency or cost budgets. We present an end-to-end industrial recommende…
ALPBench: A Benchmark for Attribution-level Long-term Personal Behavior Understanding
Lu Ren, Junda She, Xinchen Luo +23
Recent advances in large language models have highlighted their potential for personalized recommendation, where accurately capturing user preferences remains a key challenge. Leve…