2 papers
cs.CR2026
Bait-and-Recover: Poisoning Internal Refusal Signals to Defend LLMs against White-Box Editing Jailbreaks
Tian Gao, Zhipeng Xie, Yuhao Wu +2
Open-weight large language models face a low-cost white-box threat from representation engineering attacks. Attackers can estimate refusal directions and search for projection-matr…
eess.AS2026
Beyond Speech: Dual-Domain SSL Fusion for Unified All-Type Audio Deepfake Detection
Cunhang Fan, Junqin Cao, Tian Gao +4
Unified all-type audio deepfake detection aims to determine whether an input clip is real or fake when its audio type may be speech, environmental sound, singing voice, or music. E…