2 papers
cs.CL2024
Eagle: Ethical Dataset Given from Real Interactions
Masahiro Kaneko, Danushka Bollegala, Timothy Baldwin
Recent studies have demonstrated that large language models (LLMs) have ethical-related problems such as social biases, lack of moral reasoning, and generation of offensive content…
cs.CL2024
The Gaps between Pre-train and Downstream Settings in Bias Evaluation and Debiasing
Masahiro Kaneko, Danushka Bollegala, Timothy Baldwin
The output tendencies of Pre-trained Language Models (PLM) vary markedly before and after Fine-Tuning (FT) due to the updates to the model parameters. These divergences in output t…