8 citations · 12 across the 4 of their papers we have counts for
4 papers
Safeguarding Large Language Models: A Survey
Yi Dong, Ronghui Mu, Yanghao Zhang +9
In the burgeoning field of Large Language Models (LLMs), developing a robust safety mechanism, colloquially known as "safeguards" or "guardrails", has become imperative to ensure t…
Direct Training Needs Regularisation: Anytime Optimal Inference Spiking Neural Network
Dengyu Wu, Yi Qi, Kaiwen Cai +3
Spiking Neural Network (SNN) is acknowledged as the next generation of Artificial Neural Network (ANN) and hold great promise in effectively processing spatial-temporal information…
TrajPAC: Towards Robustness Verification of Pedestrian Trajectory Prediction Models
Liang Zhang, Nathaniel Xu, Pengfei Yang +3
Robust pedestrian trajectory forecasting is crucial to developing safe autonomous vehicles. Although previous works have studied adversarial robustness in the context of trajectory…
Randomized Adversarial Training via Taylor Expansion
Gaojie Jin, Xinping Yi, Dengyu Wu +2
In recent years, there has been an explosion of research into developing more robust deep neural networks against adversarial examples. Adversarial training appears as one of the m…