6 citations · 17 across the 13 of their papers we have counts for
Showing cs.CRShow all
2 papers · 1 filter
cs.CR2025
Deep Research Agents Brings Deeper Harm
Shuo Chen, Zonggen Li, Xingyu Jin +8
We reveal that Deep Research (DR) agents systematically expose safety risks: simply submitting harmful queries that a standalone LLM would reject outright can elicit detailed and d…
cs.CR2025
Strong but Brittle: Simple Attacks Subvert Reasoning-based Safety Guardrails
Shuo Chen, Zhen Han, Haokun Chen +6
Open-weight Large Reasoning Models (LRMs) are approaching the capabilities of their frontier counterparts but pose significant safety concerns, as they are difficult to patch or mo…