9 papers
Orchard: An Open-Source Agentic Modeling Framework
Baolin Peng, Wenlin Yao, Qianhui Wu +11
Agentic modeling aims to transform LLMs into autonomous agents capable of solving complex tasks through planning, reasoning, tool use, and multi-turn interaction with external envi…
STAR: Failure-Aware Markovian Routing for Multi-Agent Spatiotemporal Reasoning
Ruiyi Yang, Lihuan Li, Hao Xue +1
Compositional spatiotemporal reasoning often requires a system to invoke multiple heterogeneous specialists, such as geometric, temporal, topological, and trajectory agents. A cent…
GeoBuildBench: A Benchmark for Interactive and Executable Geometry Construction from Natural Language
Jinwoong Kim, Rui Yang, Huishuai Zhang
We introduce GeoBuildBench, a benchmark designed to evaluate whether large language models and multimodal agents can ground informal natural-language plane geometry problems into e…
AlphaGRPO: Unlocking Self-Reflective Multimodal Generation in UMMs via Decompositional Verifiable Reward
Runhui Huang, Jie Wu, Rui Yang +2
In this paper, we propose AlphaGRPO, a novel framework that applies Group Relative Policy Optimization (GRPO) to AR-Diffusion Unified Multimodal Models (UMMs) to enhance multimodal…
TrajPrism: A Multi-Task Benchmark for Language-Grounded Urban Trajectory Understanding
Lihuan Li, Wilson Wongso, Baiyu Chen +6
Urban mobility is naturally expressed both as trajectories in space and as natural-language descriptions of travel intent, constraints, and preferences. However, prior work rarely…
MAGE: Multi-Agent Self-Evolution with Co-Evolutionary Knowledge Graphs
Ruiyi Yang, Zechen Li, Hao Xue +2
Self-evolving language-model agents must decide what to learn next and how to preserve what they have learned across iterations. Existing systems typically carry this cross-iterati…