activity
20242026
collaborators
Showing cs.AIShow all

5 papers · 1 filter

cs.AI2025

Evaluating Generalization Capabilities of LLM-Based Agents in Mixed-Motive Scenarios Using Concordia

Chandler Smith, Marwa Abdulhai, Manfred Diaz +83

Large Language Model (LLM) agents have demonstrated impressive capabilities for social interaction and are increasingly being deployed in situations where they might engage with bo…

cs.AI2025

Psychometric Personality Shaping Modulates Capabilities and Safety in Language Models

Stephen Fitz, Peter Romero, Steven Basart +2

Large Language Models increasingly mediate high-stakes interactions, intensifying research on their capabilities and safety. While recent work has shown that LLMs exhibit consisten…

cs.AI2025

Beyond the high score: Prosocial ability profiles of multi-agent populations

Marko Tesic, Yue Zhao, Joel Z. Leibo +2

The development and evaluation of social capabilities in AI agents require complex environments where competitive and cooperative behaviours naturally emerge. While game-theoretic…

cs.AI2025

Identifying, Evaluating, and Mitigating Risks of AI Thought Partnerships

Kerem Oktar, Katherine M. Collins, Jose Hernandez-Orallo +4

Artificial Intelligence (AI) systems have historically been used as tools that execute narrowly defined tasks. Yet recent advances in AI have unlocked possibilities for a new class…

cs.AI2024

Conversational Complexity for Assessing Risk in Large Language Models

John Burden, Manuel Cebrian, Jose Hernandez-Orallo

Large Language Models (LLMs) present a dual-use dilemma: they enable beneficial applications while harboring potential for harm, particularly through conversational interactions. D…