activity
20242026
collaborators

6 papers

math.NT2026

Generalizations of the Christoffel-Darboux formula and congruences involving Apéry-like numbers

Zhi-Hong Sun

In this paper, we first extend the Christoffel-Darboux formula for orthogonal polynomials to general three-term recurrence sequences, and then investigate the identities and congru…

cs.AI2026

Stop Wandering, Find the Keys: LLMs Discriminate Key States for Efficient Multi-Agent Exploration

Yun Qu, Boyuan Wang, Yuhang Jiang +7

With expansive state-action spaces, efficient multi-agent exploration remains a longstanding challenge in reinforcement learning. Although pursuing novelty, diversity, or uncertain…

cs.MA2026

ATOM: Instantiating Budget-Controllable Multi-Agent Collaboration via Nucleus-Electron Hierarchy

Xinkui Zhao, Sai Liu, Yifan Zhang +6

Large Language Model (LLM)-based multi-agent systems rely on optimized collaboration topologies to balance performance and communication costs. However, current methods struggle wi…

cs.AI2026

Proximity-Based Multi-Turn Optimization: Practical Credit Assignment for LLM Agent Training

Yangyi Fang, Jiaye Lin, Xiaoliang Fu +4

Multi-turn LLM agents are becoming pivotal to production systems, spanning customer service automation, e-commerce assistance, and interactive task management, where accurately dis…

cs.LG2025

Latent Reward: LLM-Empowered Credit Assignment in Episodic Reinforcement Learning

Yun Qu, Yuhang Jiang, Boyuan Wang +4

Reinforcement learning (RL) often encounters delayed and sparse feedback in real-world applications, even with only episodic rewards. Previous approaches have made some progress in…

cs.CL2024

Falcon-UI: Understanding GUI Before Following User Instructions

Huawen Shen, Chang Liu, Gengluo Li +4

Pursuing human-like interaction for Graphical User Interface (GUI) agents requires understanding the GUI context and following user instructions. However, existing works typically…