collaborators

9 papers

cs.LG2026

MERA: Model Evolution and Routing with Skill Adaptation for Agentic Systems at Scale

Yuhang Yao, Zeyu Wang, Wanyi Chen +8

LLM agents execute heterogeneous sequences of model calls within a single task: some invocations require careful reasoning, while others are structured steps such as formatting or…

cs.LG2026

CARE: Context-Aware Ranking Evolution with Executable Scoring Programs for Budgeted Reaction Optimization

Guanyu Liu, Weiyi Kong, Chao Tang +5

High-throughput experimentation can evaluate many reaction conditions, yet combinatorial condition spaces still exceed the available experiment budget. This makes experiment select…

cs.CL2026

See, Infer, Intervene: Proactive World Modeling for Goal-Oriented Social Intelligence

Honghui Zhang, Chenmeinian Guo, Yichen Yu +7

Multimodal retail agents should not only recognize what a customer is doing, but also decide whether and how to assist before an explicit request is made. We study this setting thr…

cs.MA2026

Organizational Control Layer: Governance Infrastructure at the Execution Boundary of LLM Agent Systems

Tianyu Shi, Yang Mo, Yiou Liu +6

LLM-based agents are increasingly deployed in workflows where generated outputs may trigger state-changing actions, such as price offers, refunds, payments, or tool calls. This cre…

cs.LG2026

TwinRouterBench: Fast Static and Live Dynamic Evaluation for Realistic Agentic LLM Routing

Pei Yang, Wanyi Chen, Tongyun Yang +14

LLM routing matters most in long-horizon applications such as coding agents, deep research systems, and computer-use agents, where a single user request triggers many model calls.…

cs.LG2026

Meta-Learning Reinforcement Learning for Crypto-Return Prediction

Junqiao Wang, Zhaoyang Guan, Guanyu Liu +7

Predicting cryptocurrency returns is notoriously difficult: price movements are driven by a fast-shifting blend of on-chain activity, news flow, and social sentiment, while labeled…