activity
20222024
most citedSemantic Mirror Jailbreak: Genetic Algorithm Based Jailbreak Prompts Against Open-source LLMs

1 citations · 1 across the 5 of their papers we have counts for

collaborators

5 papers

cs.CL20241 cited

Semantic Mirror Jailbreak: Genetic Algorithm Based Jailbreak Prompts Against Open-source LLMs

Xiaoxia Li, Siyuan Liang, Jiyi Zhang +3

Large Language Models (LLMs), used in creative writing, code generation, and translation, generate text based on input sequences but are vulnerable to jailbreak attacks, where craf…

cs.LG2024

Domain Bridge: Generative model-based domain forensic for black-box models

Jiyi Zhang, Han Fang, Ee-Chien Chang

In forensic investigations of machine learning models, techniques that determine a model's data domain play an essential role, with prior work relying on large-scale corpora like I…

cs.LG2023

Adaptive Attractors: A Defense Strategy against ML Adversarial Collusion Attacks

Jiyi Zhang, Han Fang, Ee-Chien Chang

In the seller-buyer setting on machine learning models, the seller generates different copies based on the original model and distributes them to different buyers, such that advers…

cs.LG2023

Finding Meaningful Distributions of ML Black-boxes under Forensic Investigation

Jiyi Zhang, Han Fang, Hwee Kuan Lee +1

Given a poorly documented neural network model, we take the perspective of a forensic investigator who wants to find out the model's data domain (e.g. whether on face images or tra…

cs.MM2022

De-END: Decoder-driven Watermarking Network

Han Fang, Zhaoyang Jia, Yupeng Qiu +3

With recent advances in machine learning, researchers are now able to solve traditional problems with new solutions. In the area of digital watermarking, deep-learning-based waterm…