activity
20242026
most citedMACCA: Offline Multi-agent Reinforcement Learning with Causal Credit Assignment

1 citations · 1 across the 1 of their papers we have counts for

collaborators

5 papers

cs.LG20261 cited

MACCA: Offline Multi-agent Reinforcement Learning with Causal Credit Assignment

Ziyan Wang, Yali Du, Yudi Zhang +2

Offline Multi-agent Reinforcement Learning (MARL) is valuable in scenarios where online interaction is impractical or risky. While independent learning in MARL offers flexibility a…

cs.CL2025

ATLaS: Agent Tuning via Learning Critical Steps

Zhixun Chen, Ming Li, Yuxuan Huang +3

Large Language Model (LLM) agents have demonstrated remarkable generalization capabilities across multi-domain tasks. Existing agent tuning approaches typically employ supervised f…

cs.CL2025

Distill Not Only Data but Also Rewards: Can Smaller Language Models Surpass Larger Ones?

Yudi Zhang, Lu Wang, Meng Fang +8

Distilling large language models (LLMs) typically involves transferring the teacher model's responses through supervised fine-tuning (SFT). However, this approach neglects the pote…

cs.AI2025

Learning to Discuss Strategically: A Case Study on One Night Ultimate Werewolf

Xuanfa Jin, Ziyan Wang, Yali Du +3

Communication is a fundamental aspect of human society, facilitating the exchange of information and beliefs among people. Despite the advancements in large language models (LLMs),…

cs.AI2024

RuAG: Learned-rule-augmented Generation for Large Language Models

Yudi Zhang, Pei Xiao, Lu Wang +11

In-context learning (ICL) and Retrieval-Augmented Generation (RAG) have gained attention for their ability to enhance LLMs' reasoning by incorporating external knowledge but suffer…