Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
Beyond Gemini-3-Pro: Revisiting LLM Routing and Aggregation at Scale
Shengji Tang, Weihao Lin, Peng Ye +9
Large Language Models (LLMs) have rapidly advanced, with Gemini-3-Pro setting a new performance milestone. In this work, we explore collective intelligence as an alternative to mon…
cs.AI2025
Wisdom of the Crowd: Reinforcement Learning from Coevolutionary Collective Feedback
Wenzhen Yuan, Shengji Tang, Weihao Lin +8
Reinforcement learning (RL) has significantly enhanced the reasoning capabilities of large language models (LLMs), but its reliance on expensive human-labeled data or complex rewar…