1 citations · 1 across the 2 of their papers we have counts for
2 papers
cs.LG2025★ 1 cited
CURE: Critical-Token-Guided Re-Concatenation for Entropy-Collapse Prevention
Qingbin Li, Rongkun Xue, Jie Wang +8
Recent advances in Reinforcement Learning with Verified Reward (RLVR) have driven the emergence of more sophisticated cognitive behaviors in large language models (LLMs), thereby e…
cs.CV2025
GThinker: Towards General Multimodal Reasoning via Cue-Guided Rethinking
Yufei Zhan, Ziheng Wu, Yousong Zhu +10
Despite notable advancements in multimodal reasoning, leading Multimodal Large Language Models (MLLMs) still underperform on vision-centric multimodal reasoning tasks in general sc…