3 papers
cs.CL2025
Geometry-Guided Adversarial Prompt Detection via Curvature and Local Intrinsic Dimension
Canaan Yung, Hanxun Huang, Christopher Leckie +1
Adversarial prompts are capable of jailbreaking frontier large language models (LLMs) and inducing undesirable behaviours, posing a significant obstacle to their safe deployment. C…
cs.CL2024
Round Trip Translation Defence against Large Language Model Jailbreaking Attacks
Canaan Yung, Hadi Mohaghegh Dolatabadi, Sarah Erfani +1
Large language models (LLMs) are susceptible to social-engineered attacks that are human-interpretable but require a high level of comprehension for LLMs to counteract. Existing de…
quant-ph2023
Clustering by Contour coreset and variational quantum eigensolver
Canaan Yung, Muhammad Usman
Recent work has proposed solving the k-means clustering problem on quantum computers via the Quantum Approximate Optimization Algorithm (QAOA) and coreset techniques. Although the…