68 citations · 68 across the 3 of their papers we have counts for
3 papers
cs.CL2025
Exploring Multi-Temperature Strategies for Token- and Rollout-Level Control in RLVR
Haomin Zhuang, Yujun Zhou, Taicheng Guo +4
Reinforcement Learning has demonstrated substantial improvements in the reasoning abilities of Large Language Models (LLMs), exhibiting significant applicability across various dom…
cs.AI2024
SceMQA: A Scientific College Entrance Level Multimodal Question Answering Benchmark
Zhenwen Liang, Kehan Guo, Gang Liu +7
The paper introduces SceMQA, a novel benchmark for scientific multimodal question answering at the college entrance level. It addresses a critical educational phase often overlooke…
cs.CL2024★ 68 cited
Large Language Model based Multi-Agents: A Survey of Progress and Challenges
Taicheng Guo, Xiuying Chen, Yaqi Wang +5
Large Language Models (LLMs) have achieved remarkable success across a wide array of tasks. Due to the impressive planning and reasoning abilities of LLMs, they have been used as a…