2 papers
cs.CL2024
Inner-Probe: Discovering Copyright-related Data Generation in LLM Architecture
Qichao Ma, Rui-Jie Zhu, Peiye Liu +8
Large Language Models (LLMs) utilize extensive knowledge databases and show powerful text generation ability. However, their reliance on high-quality copyrighted datasets raises co…
cs.LG2024
The Exploration-Exploitation Dilemma Revisited: An Entropy Perspective
Renye Yan, Yaozhong Gan, You Wu +4
The imbalance of exploration and exploitation has long been a significant challenge in reinforcement learning. In policy optimization, excessive reliance on exploration reduces lea…