activity
20242026
collaborators

8 papers

cs.DL2026

Lacuna: A Research Map for Machine Learning

Martin Weiss, Miles Q. Li, Alejandro H. Artiles +4

Lacuna is a research map for machine learning that uses LLMs to turn papers and scholarly metadata into markdown summaries, concept elements, research directions, and research prop…

cs.AI2026

A Benchmark for Evaluating Outcome-Driven Constraint Violations in Autonomous AI Agents

Miles Q. Li, Benjamin C. M. Fung, Martin Weiss +3

As autonomous AI agents are increasingly deployed in high-stakes environments, ensuring their safety and alignment with human values is becoming a practical deployment concern. Cur…

cs.AI2026

Adaptive Prompt Embedding Optimization for LLM Jailbreaking

Miles Q. Li, Benjamin C. M. Fung, Boyang Li +2

Existing white-box jailbreak attacks against aligned LLMs typically append discrete adversarial suffixes to the user prompt, which visibly alters the prompt and operates in a combi…

cs.CV2026

LLM-as-Judge Framework for Evaluating Tone-Induced Hallucination in Vision-Language Models

Zhiyuan Jiang, Weihao Hong, Xinlei Guan +8

Vision-Language Models (VLMs) are increasingly deployed in settings where reliable visual grounding carries operational consequences, yet their behavior under progressively coerciv…

cs.CY2026

Taxonomy and Consistency Analysis of Safety Benchmarks for AI Agents

Miles Q. Li, Benjamin C. M. Fung, Boyang Li +2

The rapid deployment of LLM-based autonomous agents has introduced safety risks that extend far beyond traditional LLM concerns, prompting a proliferation of safety benchmarks sinc…

cs.CR2025

Security Concerns for Large Language Models: A Survey

Miles Q. Li, Benjamin C. M. Fung

Large Language Models (LLMs) such as ChatGPT and its competitors have caused a revolution in natural language processing, but their capabilities also introduce new security vulnera…