1 citations · 3 across the 18 of their papers we have counts for
Showing cs.LGShow all
3 papers · 1 filter
cs.LG2026
Navigate the Unknown: Enhancing LLM Reasoning with Intrinsic Motivation Guided Exploration
Jingtong Gao, Ling Pan, Yejing Wang +6
Reinforcement Learning (RL) has become a key approach for enhancing the reasoning capabilities of large language models. However, prevalent RL approaches like proximal policy optim…
cs.LG2026
Attention Needs to Focus: A Unified Perspective on Attention Allocation
Zichuan Fu, Wentao Song, Guojing Li +6
The Transformer architecture, a cornerstone of modern Large Language Models (LLMs), has achieved extraordinary success in sequence modeling, primarily due to its attention mechanis…
cs.LG2025
Generative Auto-Bidding with Value-Guided Explorations
Jingtong Gao, Yewen Li, Shuai Mao +8
Auto-bidding, with its strong capability to optimize bidding decisions within dynamic and competitive online environments, has become a pivotal strategy for advertising platforms.…