1 citations · 1 across the 2 of their papers we have counts for
2 papers
cs.LG2025
Adapting to Evolving Adversaries with Regularized Continual Robust Training
Sihui Dai, Christian Cianfarani, Arjun Bhagoji +2
Robust training methods typically defend against specific attack types, such as Lp attacks with fixed budgets, and rarely account for the fact that defenders may encounter new atta…
cs.CL2024★ 1 cited
Beyond Performance: Quantifying and Mitigating Label Bias in LLMs
Yuval Reif, Roy Schwartz
Large language models (LLMs) have shown remarkable adaptability to diverse tasks, by leveraging context prompts containing instructions, or minimal input-output examples. However,…