20 citations · 25 across the 12 of their papers we have counts for
14 papers
From Memorization to Absorption: Mixed-Policy RL for Continual Knowledge Injection
Zhibo Hou, Fan Zhao, Zhiyu An +1
Continual knowledge injection is essential for keeping large language models up-to-date in a fast-evolving world. Existing methods rely on supervised fine-tuning (SFT), which memor…
When Less Latent Leads to Better Relay: Information-Preserving Compression for Latent Multi-Agent LLM Collaboration
Yiping Li, Zhiyu An, Wan Du
Multi-agent LLM systems are moving beyond discrete-token messages toward richer relays that preserve internal state. Recent work such as LatentMAS transmits full key-value (KV) cac…
Open Problems in Differentiable Social Choice: Learning Mechanisms, Decisions, and Alignment
Zhiyu An, Wan Du
Social choice has become a foundational component of modern machine learning systems. From auctions and resource allocation to the alignment of large generative models, machine lea…
Differential Voting: Loss Functions For Axiomatically Diverse Aggregation of Heterogeneous Preferences
Zhiyu An, Duaa Nakshbandi, Wan Du
Reinforcement learning from human feedback (RLHF) implicitly aggregates heterogeneous human preferences into a single utility function, even though the underlying utilities of the…
DIML: Differentiable Inverse Mechanism Learning from Behaviors of Multi-Agent Learning Trajectories
Zhiyu An, Wan Du
We study inverse mechanism learning: recovering an unknown incentive-generating mechanism from observed strategic interaction traces of self-interested learning agents. Unlike inve…
Representational Homomorphism Predicts and Improves Compositional Generalization In Transformer Language Model
Zhiyu An, Wan Du
Compositional generalization-the ability to interpret novel combinations of familiar components-remains a persistent challenge for neural networks. Behavioral evaluations reveal \e…