Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
The Invisible Leash: Why RLVR May or May Not Escape Its Origin
Fang Wu, Weihao Xuan, Ximing Lu +4
Recent advances highlight Reinforcement Learning with Verifiable Rewards (RLVR) as a promising method for enhancing LLMs' capabilities. However, it remains unclear whether the curr…
cs.LG2025
Is Pre-training Applicable to the Decoder for Dense Prediction?
Chao Ning, Wanshui Gan, Weihao Xuan +1
Pre-trained encoders are widely employed in dense prediction tasks for their capability to effectively extract visual features from images. The decoder subsequently processes these…