activity
20242026
most citedMARLP: Time-series Forecasting Control for Agricultural Managed Aquifer Recharge

20 citations · 25 across the 12 of their papers we have counts for

collaborators

14 papers

cs.CL2026

From Memorization to Absorption: Mixed-Policy RL for Continual Knowledge Injection

Zhibo Hou, Fan Zhao, Zhiyu An +1

Continual knowledge injection is essential for keeping large language models up-to-date in a fast-evolving world. Existing methods rely on supervised fine-tuning (SFT), which memor…

cs.LG2026

When Less Latent Leads to Better Relay: Information-Preserving Compression for Latent Multi-Agent LLM Collaboration

Yiping Li, Zhiyu An, Wan Du

Multi-agent LLM systems are moving beyond discrete-token messages toward richer relays that preserve internal state. Recent work such as LatentMAS transmits full key-value (KV) cac…

cs.AI2026

Open Problems in Differentiable Social Choice: Learning Mechanisms, Decisions, and Alignment

Zhiyu An, Wan Du

Social choice has become a foundational component of modern machine learning systems. From auctions and resource allocation to the alignment of large generative models, machine lea…

cs.GT2026

Differential Voting: Loss Functions For Axiomatically Diverse Aggregation of Heterogeneous Preferences

Zhiyu An, Duaa Nakshbandi, Wan Du

Reinforcement learning from human feedback (RLHF) implicitly aggregates heterogeneous human preferences into a single utility function, even though the underlying utilities of the…

cs.AI2026

DIML: Differentiable Inverse Mechanism Learning from Behaviors of Multi-Agent Learning Trajectories

Zhiyu An, Wan Du

We study inverse mechanism learning: recovering an unknown incentive-generating mechanism from observed strategic interaction traces of self-interested learning agents. Unlike inve…

cs.LG2026

Representational Homomorphism Predicts and Improves Compositional Generalization In Transformer Language Model

Zhiyu An, Wan Du

Compositional generalization-the ability to interpret novel combinations of familiar components-remains a persistent challenge for neural networks. Behavioral evaluations reveal \e…