1 citations · 1 across the 2 of their papers we have counts for
1 paper · 1 filter
Ching-An Cheng, Andrey Kolobov, Dipendra Misra +2
We introduce a new benchmark, LLF-Bench (Learning from Language Feedback Benchmark; pronounced as "elf-bench"), to evaluate the ability of AI agents to interactively learn from nat…