Showing cs.AIShow all
2 papers · 1 filter
cs.AI2025
BALROG: Benchmarking Agentic LLM and VLM Reasoning On Games
Davide Paglieri, BartÅomiej CupiaÅ, Samuel Coward +10
Large Language Models (LLMs) and Vision Language Models (VLMs) possess extensive knowledge and exhibit promising reasoning abilities, however, they still struggle to perform well i…
cs.AI2024
diff History for Neural Language Agents
Ulyana Piterbarg, Lerrel Pinto, Rob Fergus
Neural Language Models (LMs) offer an exciting solution for general-purpose embodied control. However, a key technical issue arises when using an LM-based controller: environment o…