Showing 2024Show all
2 papers · 1 filter
cs.CL2024
Delving into the Reversal Curse: How Far Can Large Language Models Generalize?
Zhengkai Lin, Zhihang Fu, Kai Liu +6
While large language models (LLMs) showcase unprecedented capabilities, they also exhibit certain inherent limitations when facing seemingly trivial tasks. A prime example is the r…
cs.CL2024
Enhancing Multiple Dimensions of Trustworthiness in LLMs via Sparse Activation Control
Yuxin Xiao, Chaoqun Wan, Yonggang Zhang +5
As the development and application of Large Language Models (LLMs) continue to advance rapidly, enhancing their trustworthiness and aligning them with human preferences has become…