4 papers · 1 filter
LongCat-Flash-Prover: Advancing Native Formal Reasoning via Agentic Tool-Integrated Reinforcement Learning
Jianing Wang, Jianfei Zhang, Qi Guo +24
We introduce LongCat-Flash-Prover, a flagship 560-billion-parameter open-source Mixture-of- Experts (MoE) model that advances Native Formal Reasoning in Lean4 through agentic tool-…
Cognitively Layered Data Synthesis for Domain Adaptation of LLMs to Space Situational Awareness
Ding Linghu, Cheng Wang, Da Fan +6
Large language models (LLMs) demonstrate exceptional performance on general-purpose tasks. however, transferring them to complex engineering domains such as space situational aware…
InstructRAG: Leveraging Retrieval-Augmented Generation on Instruction Graphs for LLM-Based Task Planning
Zheng Wang, Shu Xian Teo, Jun Jie Chew +1
Recent advancements in large language models (LLMs) have enabled their use as agents for planning complex tasks. Existing methods typically rely on a thought-action-observation (TA…
MASTER: A Multi-Agent System with LLM Specialized MCTS
Bingzheng Gan, Yufan Zhao, Tianyi Zhang +5
Large Language Models (LLM) are increasingly being explored for problem-solving tasks. However, their strategic planning capability is often viewed with skepticism. Recent studies…