3 papers
cs.AI2025
PillagerBench: Benchmarking LLM-Based Agents in Competitive Minecraft Team Environments
Olivier Schipper, Yudi Zhang, Yali Du +2
LLM-based agents have shown promise in various cooperative and strategic reasoning tasks, but their effectiveness in competitive multi-agent environments remains underexplored. To…
cs.LG2025
Learning Instruction-Following Policies through Open-Ended Instruction Relabeling with Large Language Models
Zhicheng Zhang, Ziyan Wang, Yali Du +1
Developing effective instruction-following policies in reinforcement learning remains challenging due to the reliance on extensive human-labeled instruction datasets and the diffic…
cs.AI2024
RuAG: Learned-rule-augmented Generation for Large Language Models
Yudi Zhang, Pei Xiao, Lu Wang +11
In-context learning (ICL) and Retrieval-Augmented Generation (RAG) have gained attention for their ability to enhance LLMs' reasoning by incorporating external knowledge but suffer…