1 citations · 1 across the 3 of their papers we have counts for
1 paper · 1 filter
Alexander Bondarenko, Denis Volk, Dmitrii Volkov +1
We demonstrate LLM agent specification gaming by instructing models to win against a chess engine. We find reasoning models like OpenAI o3 and DeepSeek R1 will often hack the bench…