Showing cs.CRShow all
2 papers · 1 filter
cs.CR2025
Pushing the Limits of Safety: A Technical Report on the ATLAS Challenge 2025
Zonghao Ying, Siyang Wu, Run Hao +44
Multimodal Large Language Models (MLLMs) have enabled transformative advancements across diverse applications but remain susceptible to safety threats, especially jailbreak attacks…
cs.CR2024
Exploring Query Efficient Data Generation towards Data-free Model Stealing in Hard Label Setting
Gaozheng Pei, Shaojie lyu, Ke Ma +3
Data-free model stealing involves replicating the functionality of a target model into a substitute model without accessing the target model's structure, parameters, or training da…