5 citations · 5 across the 2 of their papers we have counts for
2 papers
cs.LG2026
Mitigating Overthinking in Large Reasoning Models via Difficulty-aware Reinforcement Learning
Qian Wan, Ziao Xu, Luona Wei +2
Large Reasoning Models (LRMs) achieve explicit chain-of-thought expansion by imitating deep thinking behaviors of humans, demonstrating excellent performance in complex task scenar…
cs.NE2022★ 5 cited
RL-GA: A Reinforcement Learning-Based Genetic Algorithm for Electromagnetic Detection Satellite Scheduling Problem
Yanjie Song, Luona Wei, Qing Yang +3
The study of electromagnetic detection satellite scheduling problem (EDSSP) has attracted attention due to the detection requirements for a large number of targets. This paper prop…