20 citations · 40 across the 10 of their papers we have counts for
Showing cs.CLShow all
2 papers · 1 filter
cs.CL2023★ 2 cited
Mitigating Shortcuts in Language Models with Soft Label Encoding
Zirui He, Huiqi Deng, Haiyan Zhao +2
Recent research has shown that large language models rely on spurious correlations in the data for natural language understanding (NLU) tasks. In this work, we aim to answer the fo…
cs.CL2023★ 20 cited
Explainability for Large Language Models: A Survey
Haiyan Zhao, Hanjie Chen, Fan Yang +6
Large language models (LLMs) have demonstrated impressive capabilities in natural language processing. However, their internal mechanisms are still unclear and this lack of transpa…