2 papers
cs.AI2025
Misalignment Bounty: Crowdsourcing AI Agent Misbehavior
Rustem Turtayev, Natalia Fedorova, Oleg Serikov +3
Advanced AI systems sometimes act in ways that differ from human intent. To gather clear, reproducible examples, we ran the Misalignment Bounty: a crowdsourced project that collect…
cs.AI2024
How to Correctly do Semantic Backpropagation on Language-based Agentic Systems
Wenyi Wang, Hisham A. Alyahya, Dylan R. Ashley +4
Language-based agentic systems have shown great promise in recent years, transitioning from solving small-scale research problems to being deployed in challenging real-world tasks.…