collaborators

8 papers

cs.LG2026

When Less Latent Leads to Better Relay: Information-Preserving Compression for Latent Multi-Agent LLM Collaboration

Yiping Li, Zhiyu An, Wan Du

Multi-agent LLM systems are moving beyond discrete-token messages toward richer relays that preserve internal state. Recent work such as LatentMAS transmits full key-value (KV) cac…

cs.LG2026

Beyond Noisy-TVs: Noise-Robust Exploration Via Learning Progress Monitoring

Zhibo Hou, Zhiyu An, Wan Du

When there exists an unlearnable source of randomness (noisy-TV) in the environment, a naively intrinsic reward driven exploring agent gets stuck at that source of randomness and f…

cs.LG2026

Representational Homomorphism Predicts and Improves Compositional Generalization In Transformer Language Model

Zhiyu An, Wan Du

Compositional generalization-the ability to interpret novel combinations of familiar components-remains a persistent challenge for neural networks. Behavioral evaluations reveal \e…

cs.AI2026

Open Problems in Differentiable Social Choice: Learning Mechanisms, Decisions, and Alignment

Zhiyu An, Wan Du

Social choice has become a foundational component of modern machine learning systems. From auctions and resource allocation to the alignment of large generative models, machine lea…

cs.GT2026

Differential Voting: Loss Functions For Axiomatically Diverse Aggregation of Heterogeneous Preferences

Zhiyu An, Duaa Nakshbandi, Wan Du

Reinforcement learning from human feedback (RLHF) implicitly aggregates heterogeneous human preferences into a single utility function, even though the underlying utilities of the…

cs.AI2026

DIML: Differentiable Inverse Mechanism Learning from Behaviors of Multi-Agent Learning Trajectories

Zhiyu An, Wan Du

We study inverse mechanism learning: recovering an unknown incentive-generating mechanism from observed strategic interaction traces of self-interested learning agents. Unlike inve…