8 citations · 8 across the 5 of their papers we have counts for
Showing cs.AIShow all
2 papers · 1 filter
cs.AI2025
RoRecomp: Enhancing Reasoning Efficiency via Rollout Response Recomposition in Reinforcement Learning
Gang Li, Yulei Qin, Xiaoyu Tan +6
Reinforcement learning with verifiable rewards (RLVR) has proven effective in eliciting complex reasoning in large language models (LLMs). However, standard RLVR training often lea…
cs.AI2025
Towards deployment-centric multimodal AI beyond vision and language
Xianyuan Liu, Jiayang Zhang, Shuo Zhou +45
Multimodal artificial intelligence (AI) integrates diverse types of data via machine learning to improve understanding, prediction, and decision-making across disciplines such as h…