8 papers
Benchmarking LLMs for Community Governance Simulation with Life-history Narratives
Xu Chen, Yuanzi Li, Lei Wang +6
Effective community governance hinges on understanding what specific residents think and need. Recent work has used large language models (LLMs) to simulate human respondents, offe…
FedMPT: Federated Multi-label Prompt Tuning of Vision-Language Models
Xucong Wang, Pengkun Wang, Zhe Zhao +3
Multi-Label Recognition (MLR) based on Vision-Language Models (VLMs) aims to leverage their pre-trained knowledge to better adapt complex recognition scenarios, thereby enhancing m…
The MiniMax-M2 Series: Mini Activations Unleashing Max Real-World Intelligence
MiniMax, :, Aili Chen +219
We introduce the MiniMax-M2 series, a family of Mixture-of-Experts language models built around the principle that mini activations can unleash maximum real-world intelligence. The…
AcademiClaw: When Students Set Challenges for AI Agents
Junjie Yu, Pengrui Lu, Weiye Si +75
Benchmarks within the OpenClaw ecosystem have thus far evaluated exclusively assistant-level tasks, leaving the academic-level capabilities of OpenClaw largely unexamined. We intro…
Partial GFlowNet: Accelerating Convergence in Large State Spaces via Strategic Partitioning
Xuan Yu, Xu Wang, Rui Zhu +2
Generative Flow Networks (GFlowNets) have shown promising potential to generate high-scoring candidates with probability proportional to their rewards. As existing GFlowNets freely…
Exploring Multiple High-Scoring Subspaces in Generative Flow Networks
Xuan Yu, Xu Wang, Rui Zhu +2
As a probabilistic sampling framework, Generative Flow Networks (GFlowNets) show strong potential for constructing complex combinatorial objects through the sequential composition…