1 citations · 1 across the 3 of their papers we have counts for
3 papers
cs.LG2026
Gradients Know What Outcomes Don't: Unlocking Reinforcement Learning for LLM Reasoning with Gradient-Aligned Rewards
Leqi Zheng, Jinbo Su, Fang Niu +8
Reinforcement learning from verifiable rewards (RLVR) drives chain-of-thought reasoning in large language models, yet its binary outcome reward cannot distinguish among correct tra…
cs.IR2026
SciLENS: RL-Driven Autonomous Agents for Scientific Localized Evidence Navigation and Synthesis
Leqi Zheng, Jinbo Su, Yuying Li +10
Scientific literature synthesis agents increasingly rely on proprietary online services, limiting reproducibility, privacy, and offline deployment. To address this challenge, we in…
cs.IR2026★ 1 cited
What Should I Cite? A RAG Benchmark for Academic Citation Prediction
Leqi Zheng, Jiajun Zhang, Canzhi Chen +13
With the rapid growth of Web-based academic publications, more and more papers are being published annually, making it increasingly difficult to find relevant prior work. Citation…