adversarial example generation 1agentic data curation 1automated red teaming 1content safety 1multimodal language models 1
From the 1 of 3 linked papers with an AI index.
3 papers
cs.AI2026
Automatic Hard Example Synthesis with Multi-Level Agentic Data Curation
Genglin Liu, Muye Zhang, Krishnamurthy Viswanathan +5
The paper introduces an automated, multi‑agent framework that creates hard adversarial examples for multimodal large language models to improve content safety, achieving a signific…
cs.LG2025
SYNAPSE-G: Bridging Large Language Models and Graph Learning for Rare Event Classification
Sasan Tavakkol, Lin Chen, Max Springer +4
Scarcity of labeled data, especially for rare events, hinders training effective machine learning models. This paper proposes SYNAPSE-G (Synthetic Augmentation for Positive Samplin…
cs.CR2024
Supporting Human Raters with the Detection of Harmful Content using Large Language Models
Kurt Thomas, Patrick Gage Kelley, David Tao +7
In this paper, we explore the feasibility of leveraging large language models (LLMs) to automate or otherwise assist human raters with identifying harmful content including hate sp…