Showing cs.CLShow all
2 papers · 1 filter
cs.CL2025
SSRL: Self-Search Reinforcement Learning
Yuchen Fan, Kaiyan Zhang, Heng Zhou +15
We investigate the potential of large language models (LLMs) to serve as efficient simulators for agentic search tasks in reinforcement learning (RL), thereby reducing dependence o…
cs.CL2025
Tracing Facts or just Copies? A critical investigation of the Competitions of Mechanisms in Large Language Models
Dante Campregher, Yanxu Chen, Sander Hoffman +1
This paper presents a reproducibility study examining how Large Language Models (LLMs) manage competing factual and counterfactual information, focusing on the role of attention he…