1 citations · 1 across the 5 of their papers we have counts for
4 papers · 1 filter
Where Detectors Fail: Closing the Tail-Domain Gap with Expert-Guided Mutual Distillation
Xuan Feng, Guihong Liu, Tianlong Gu +5
Multimodal fake news detectors often generalize poorly across domains because they learn to trust unreliable evidence: domain-specific shortcuts amplified by imbalanced data and se…
Self-Debias: Self-correcting for Debiasing Large Language Models
Xuan Feng, Shuai Zhao, Luwei Xiao +2
Although Large Language Models (LLMs) demonstrate remarkable reasoning capabilities, inherent social biases often cascade throughout the Chain-of-Thought (CoT) process, leading to…
C2PO: Diagnosing and Disentangling Bias Shortcuts in LLMs
Xuan Feng, Bo An, Tianlong Gu +4
Bias in Large Language Models (LLMs) poses significant risks to trustworthiness, manifesting primarily as stereotypical biases (e.g., gender or racial stereotypes) and structural b…
Learning from Mistakes: Self-correct Adversarial Training for Chinese Unnatural Text Correction
Xuan Feng, Tianlong Gu, Xiaoli Liu +1
Unnatural text correction aims to automatically detect and correct spelling errors or adversarial perturbation errors in sentences. Existing methods typically rely on fine-tuning o…