Showing cs.CLShow all
2 papers · 1 filter
cs.CL2025
C2PO: Diagnosing and Disentangling Bias Shortcuts in LLMs
Xuan Feng, Bo An, Tianlong Gu +4
Bias in Large Language Models (LLMs) poses significant risks to trustworthiness, manifesting primarily as stereotypical biases (e.g., gender or racial stereotypes) and structural b…
cs.CL2024
Learning from Mistakes: Self-correct Adversarial Training for Chinese Unnatural Text Correction
Xuan Feng, Tianlong Gu, Xiaoli Liu +1
Unnatural text correction aims to automatically detect and correct spelling errors or adversarial perturbation errors in sentences. Existing methods typically rely on fine-tuning o…