4 papers · 1 filter
Layer barriers for colour-biased tight Hamilton cycles
Zijian Deng, Qinfei Tang, Caihong Yang
We construct a family of layer barriers for colour-biased tight Hamilton cycles in uniform hypergraphs. For every and every , we give a red--blue col…
ACPO: Agent-Chained Policy Optimization for Multi-Agent Reinforcement Learning
Daiki E. Matsunaga, Junho Na, Tri Wahyu Guntara +4
Cooperative tasks in Multi-Agent Reinforcement Learning (MARL) require agents to collectively maximize a shared return. Under the Centralized Training with Decentralized Execution…
Counterfactual Residual Data Augmentation for Regression
Hossein Mohebbi, Oliver Schulte, Ke Li +1
Data-driven modeling in real-world regression tasks often suffers from limited training samples, high collection costs, and noisy observations. Inspired by the impact of data augme…
Talk, Judge, Cooperate: Gossip-Driven Indirect Reciprocity in Self-Interested LLM Agents
Shuhui Zhu, Yue Lin, Shriya Kaistha +5
Indirect reciprocity, which means helping those who have helped others, is difficult to sustain among decentralized, self-interested LLM agents without reliable reputation systems.…