11 papers
Layer barriers for colour-biased tight Hamilton cycles
Zijian Deng, Qinfei Tang, Caihong Yang
We construct a family of layer barriers for colour-biased tight Hamilton cycles in uniform hypergraphs. For every and every , we give a red--blue col…
ACPO: Agent-Chained Policy Optimization for Multi-Agent Reinforcement Learning
Daiki E. Matsunaga, Junho Na, Tri Wahyu Guntara +4
Cooperative tasks in Multi-Agent Reinforcement Learning (MARL) require agents to collectively maximize a shared return. Under the Centralized Training with Decentralized Execution…
Counterfactual Residual Data Augmentation for Regression
Hossein Mohebbi, Oliver Schulte, Ke Li +1
Data-driven modeling in real-world regression tasks often suffers from limited training samples, high collection costs, and noisy observations. Inspired by the impact of data augme…
Talk, Judge, Cooperate: Gossip-Driven Indirect Reciprocity in Self-Interested LLM Agents
Shuhui Zhu, Yue Lin, Shriya Kaistha +5
Indirect reciprocity, which means helping those who have helped others, is difficult to sustain among decentralized, self-interested LLM agents without reliable reputation systems.…
Policy-Conditioned Policies for Multi-Agent Task Solving
Yue Lin, Shuhui Zhu, Wenhao Li +5
In multi-agent tasks, the central challenge lies in the dynamic adaptation of strategies. However, directly conditioning on opponents' strategies is intractable in the prevalent de…
Image-POSER: Reflective RL for Multi-Expert Image Generation and Editing
Hossein Mohebbi, Mohammed Abdulrahman, Yanting Miao +2
Recent advances in text-to-image generation have produced strong single-shot models, yet no individual system reliably executes the long, compositional prompts typical of creative…