6 citations · 6 across the 17 of their papers we have counts for
1 paper · 1 filter
Víctor Gallego
Can large language model agents discover hidden safety objectives through experience alone? We introduce EPO-Safe (Experiential Prompt Optimization for Safe Agents), a framework wh…