Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
Logics-STEM: Empowering LLM Reasoning via Failure-Driven Post-Training and Document Knowledge Enhancement
Mingyu Xu, Cheng Fang, Keyue Jiang +16
We present Logics-STEM, a state-of-the-art reasoning model fine-tuned on Logics-STEM-SFT-Dataset, a high-quality and diverse dataset at 10M scale that represents one of the largest…
cs.AI2025
On the Importance of Task Complexity in Evaluating LLM-Based Multi-Agent Systems
Bohan Tang, Huidong Liang, Keyue Jiang +1
Large language model multi-agent systems (LLM-MAS) offer a promising paradigm for harnessing collective intelligence to achieve more advanced forms of AI behaviour. While recent st…