most citedLatent Jailbreak: A Benchmark for Evaluating Text Safety and Output Robustness of Large Language Models

11 citations · 13 across the 4 of their papers we have counts for

collaborators

4 papers