Showing cs.CLShow all
2 papers · 1 filter
cs.CL2025
C2PO: Diagnosing and Disentangling Bias Shortcuts in LLMs
Xuan Feng, Bo An, Tianlong Gu +4
Bias in Large Language Models (LLMs) poses significant risks to trustworthiness, manifesting primarily as stereotypical biases (e.g., gender or racial stereotypes) and structural b…
cs.CL2024
CHEAT: A Large-scale Dataset for Detecting ChatGPT-writtEn AbsTracts
Peipeng Yu, Jiahan Chen, Xuan Feng +1
The powerful ability of ChatGPT has caused widespread concern in the academic community. Malicious users could synthesize dummy academic content through ChatGPT, which is extremely…