most citedJailbreak Large Vision-Language Models Through Multi-Modal Linkage

1 citations · 2 across the 5 of their papers we have counts for

collaborators

5 papers

cs.CV2025

NP-LoRA: Null Space Projection for Subject-Style LoRA Fusion

Chuheng Chen, Xiaofei Zhou, Geyuan Zhang +1

Low-Rank Adaptation (LoRA) fusion enables the composition of subject and style representations for controllable generation without retraining. However, existing approaches primaril…

cs.CV2025

LaRe: Latent Refocusing for Multimodal Reasoning

Jizheng Ma, Xiaofei Zhou, Geyuan Zhang +2

Chain of Thought (CoT) reasoning enhances logical performance by decomposing complex tasks, yet its multimodal extension faces a trade-off. The prevailing Thinking with Images para…

cs.CR2025

AICrypto: Evaluating Cryptography Capabilities of Large Language Models

Yu Wang, Yijian Liu, Liheng Ji +11

We build \textbf{AICrypto}, a comprehensive benchmark designed to evaluate the cryptography capabilities of large language models (LLMs). The benchmark comprises 135 multiple-choic…

cs.CV2024

Jailbreak Large Vision-Language Models Through Multi-Modal Linkage

Yu Wang, Xiaofei Zhou, Yichen Wang +2

With the significant advancement of Large Vision-Language Models (VLMs), concerns about their potential misuse and abuse have grown rapidly. Previous studies have highlighted VLMs'…

cs.CL2024

HUT: A More Computation Efficient Fine-Tuning Method With Hadamard Updated Transformation

Geyuan Zhang, Xiaofei Zhou, Chuheng Chen

Fine-tuning pre-trained language models for downstream tasks has achieved impressive results in NLP. However, fine-tuning all parameters becomes impractical due to the rapidly incr…