3 papers
cs.CR2025
Defending against Indirect Prompt Injection by Instruction Detection
Tongyu Wen, Chenglong Wang, Xiyuan Yang +5
The integration of Large Language Models (LLMs) with external sources is becoming increasingly common, with Retrieval-Augmented Generation (RAG) being a prominent example. However,…
cs.CR2025
Evidencing Unauthorized Training Data from AI Generated Content using Information Isotopes
Qi Tao, Yin Jinhua, Cai Dongqi +10
In light of scaling laws, many AI institutions are intensifying efforts to construct advanced AIs on extensive collections of high-quality human data. However, in a rush to stay co…
cs.CY2025
Variance reduction in output from generative AI
Yu Xie, Yueqi Xie
Generative AI models, such as ChatGPT, will increasingly replace humans in producing output for a variety of important tasks. While much prior work has mostly focused on the improv…