3 citations · 3 across the 3 of their papers we have counts for
3 papers
CODA: Coordination via On-Policy Diffusion for Multi-Agent Offline Reinforcement Learning
Marcel Hedman, Kale-ab Abebe Tessera, Juan Claude Formanek +5
Offline multi-agent reinforcement learning (MARL) enables policy learning from fixed datasets, but is prone to coordination failure: agents trained on static, off-policy data conve…
LUCID: LLM-Generated Utterances for Complex and Interesting Dialogues
Joe Stacey, Jianpeng Cheng, John Torr +5
Spurred by recent advances in Large Language Models (LLMs), virtual assistants are poised to take a leap forward in terms of their dialogue capabilities. Yet a major bottleneck to…
Intelligent Assistant Language Understanding On Device
Cecilia Aas, Hisham Abdelsalam, Irina Belousova +20
It has recently become feasible to run personal digital assistants on phones and other personal devices. In this paper we describe a design for a natural language understanding sys…