2 papers
cs.CR2025
SandboxEval: Towards Securing Test Environment for Untrusted Code
Rafiqul Rabin, Jesse Hostetler, Sean McGregor +2
While large language models (LLMs) are powerful assistants in programming tasks, they may also produce malicious code. Testing LLM-generated code therefore poses significant risks…
cs.CR2025
Malicious and Unintentional Disclosure Risks in Large Language Models for Code Generation
Rafiqul Rabin, Sean McGregor, Nick Judd
This paper explores the risk that a large language model (LLM) trained for code generation on data mined from software repositories will generate content that discloses sensitive i…