1 paper
Mingqian Feng, Xiaodong Liu, Weiwei Yang +4
Multi-turn jailbreaks capture the real threat model for safety-aligned chatbots, where single-turn attacks are merely a special case. Yet existing approaches break under exploratio…