11 papers · 1 filter
Aligning Large Vision-Language Models at Test Time: A Trajectory-Guided Structured Sampling Approach
Tianbao Jiang, Weicong Ni, Gerard de Melo +1
Post-training reinforcement learning (RL) algorithms are commonly used to align large vision-language models (LVLMs) with human intent and the requirements of visual reasoning task…
From Construction to Injection: Edit-Based Fingerprints for Large Language Models
Yue Li, Xin Yi, Dongsheng Shi +3
Reliable model fingerprints are essential for protecting large language models (LLMs) against unauthorized redistribution and commercial misuse. In black-box deployment, verificati…
SURGENT: A Surgical Multi-Agent Assistance System Across the Perioperative Workflow
Dongsheng Shi, Yue Li, Xin Yi +3
The intricate nature of modern surgical care necessitates intelligent systems that can synthesize extensive patient records, support collaborative decision-making, and provide tran…
Unified Defense for Large Language Models against Jailbreak and Fine-Tuning Attacks in Education
Xin Yi, Yue Li, Dongsheng Shi +3
Large Language Models (LLMs) are increasingly integrated into educational applications. However, they remain vulnerable to jailbreak and fine-tuning attacks, which can compromise s…
Unified attacks to large language model watermarks: spoofing and scrubbing in unauthorized knowledge distillation
Xin Yi, Yue Li, Shunfan Zheng +3
Watermarking has emerged as a critical technique for combating misinformation and protecting intellectual property in large language models (LLMs). A recent discovery, termed water…
Hierarchical Safety Realignment: Lightweight Restoration of Safety in Pruned Large Vision-Language Models
Yue Li, Xin Yi, Dongsheng Shi +3
With the increasing size of Large Vision-Language Models (LVLMs), network pruning techniques aimed at compressing models for deployment in resource-constrained environments have ga…