adversarial robustness 1collaborative fine-tuning 1embedding alignment 1model consolidation 1vision-language models 1
From the 1 of 11 linked papers with an AI index.
Showing cs.CRShow all
2 papers · 1 filter
cs.CR2026
Adversarial Attacks for Good: A Survey of Proactive Protection across the Visual Content Lifecycle
Jiaming Zhang, Boyang Chen, Zherui Li +14
Once visual content enters an AI pipeline, its owner often retains little technical control over how it is used. Legal and regulatory remedies can address misuse, but many technica…
cs.CR2026
Adaptive Probe-based Steering for Robust LLM Jailbreaking
Junxi Chen, Junhao Dong, Xiaohua Xie
Recent work has demonstrated the potential of contrastive steering for jailbreaking Large Language Models (LLMs). However, existing methods rely on limited and inherently biased co…