2 papers
cs.AI2024
SPO: Sequential Monte Carlo Policy Optimisation
Matthew V Macfarlane, Edan Toledo, Donal Byrne +2
Leveraging planning during learning and decision-making is central to the long-term development of intelligent agents. Recent works have successfully combined tree-based search met…
cs.CL2024
Should we be going MAD? A Look at Multi-Agent Debate Strategies for LLMs
Andries Smit, Paul Duckworth, Nathan Grinsztajn +2
Recent advancements in large language models (LLMs) underscore their potential for responding to inquiries in various domains. However, ensuring that generative agents provide accu…