4 papers
Learning When to Cooperate Under Heterogeneous Goals
Max Taylor-Davies, Neil Bramley, Christopher G. Lucas
A significant element of human cooperative intelligence lies in our ability to identify opportunities for fruitful collaboration; and conversely to recognise when the task at hand…
Partner Modelling Emerges in Recurrent Agents (But Only When It Matters)
Ruaridh Mon-Williams, Max Taylor-Davies, Elizabeth Mieczkowski +5
Humans are remarkably adept at collaboration, able to infer the strengths and weaknesses of new partners in order to work successfully towards shared goals. To build AI systems wit…
Is Feedback All You Need? Leveraging Natural Language Feedback in Goal-Conditioned Reinforcement Learning
Sabrina McCallum, Max Taylor-Davies, Stefano V. Albrecht +1
Despite numerous successes, the field of reinforcement learning (RL) remains far from matching the impressive generalisation power of human behaviour learning. One possible way to…
Balancing utility and cognitive cost in social representation
Max Taylor-Davies, Christopher G. Lucas
To successfully navigate its environment, an agent must construct and maintain representations of the other agents that it encounters. Such representations are useful for many task…