adversarial attacks 1denoising smoothing 1large language models 1robustness certification 1semantic clustering 1
From the 1 of 3 linked papers with an AI index.
3 papers
cs.LG2026
CluCERT: Certifying LLM Robustness via Clustering-Guided Denoising Smoothing
Zixia Wang, Gaojie Jin, Jia Hu +1
The paper presents CluCERT, a framework that uses clustering-guided denoising smoothing to certify the robustness of large language models against adversarial synonym substitutions…
cs.LG2026
OTora: A Unified Red Teaming Framework for Reasoning-Level Denial-of-Service in LLM Agents
Xinyu Li, Ronghui Mu, Lin Li +2
Large Language Models (LLMs) are increasingly deployed as autonomous agents that execute tool-augmented, multi-step tasks, where latency is a critical factor for real-world applica…
cs.AI2025
Safety of Embodied Navigation: A Survey
Zixia Wang, Jia Hu, Ronghui Mu
As large language models (LLMs) continue to advance and gain influence, the development of embodied AI has accelerated, drawing significant attention, particularly in navigation sc…