8 papers
Lacuna: A Research Map for Machine Learning
Martin Weiss, Miles Q. Li, Alejandro H. Artiles +4
Lacuna is a research map for machine learning that uses LLMs to turn papers and scholarly metadata into markdown summaries, concept elements, research directions, and research prop…
A Benchmark for Evaluating Outcome-Driven Constraint Violations in Autonomous AI Agents
Miles Q. Li, Benjamin C. M. Fung, Martin Weiss +3
As autonomous AI agents are increasingly deployed in high-stakes environments, ensuring their safety and alignment with human values is becoming a practical deployment concern. Cur…
Adaptive Prompt Embedding Optimization for LLM Jailbreaking
Miles Q. Li, Benjamin C. M. Fung, Boyang Li +2
Existing white-box jailbreak attacks against aligned LLMs typically append discrete adversarial suffixes to the user prompt, which visibly alters the prompt and operates in a combi…
LLM-as-Judge Framework for Evaluating Tone-Induced Hallucination in Vision-Language Models
Zhiyuan Jiang, Weihao Hong, Xinlei Guan +8
Vision-Language Models (VLMs) are increasingly deployed in settings where reliable visual grounding carries operational consequences, yet their behavior under progressively coerciv…
Taxonomy and Consistency Analysis of Safety Benchmarks for AI Agents
Miles Q. Li, Benjamin C. M. Fung, Boyang Li +2
The rapid deployment of LLM-based autonomous agents has introduced safety risks that extend far beyond traditional LLM concerns, prompting a proliferation of safety benchmarks sinc…
Security Concerns for Large Language Models: A Survey
Miles Q. Li, Benjamin C. M. Fung
Large Language Models (LLMs) such as ChatGPT and its competitors have caused a revolution in natural language processing, but their capabilities also introduce new security vulnera…