1 citations · 1 across the 1 of their papers we have counts for
2 papers
cs.AI2025
Misalignment Bounty: Crowdsourcing AI Agent Misbehavior
Rustem Turtayev, Natalia Fedorova, Oleg Serikov +3
Advanced AI systems sometimes act in ways that differ from human intent. To gather clear, reproducible examples, we ran the Misalignment Bounty: a crowdsourced project that collect…
cs.CR2024★ 1 cited
Hacking CTFs with Plain Agents
Rustem Turtayev, Artem Petrov, Dmitrii Volkov +1
We saturate a high-school-level hacking benchmark with plain LLM agent design. Concretely, we obtain 95% performance on InterCode-CTF, a popular offensive security benchmark, using…