Showing cs.CLShow all
2 papers · 1 filter
cs.CL2024
Interpreting Key Mechanisms of Factual Recall in Transformer-Based Language Models
Ang Lv, Yuhan Chen, Kaiyi Zhang +5
In this paper, we delve into several mechanisms employed by Transformer-based language models (LLMs) for factual recall tasks. We outline a pipeline consisting of three major steps…
cs.CL2023
Baichuan 2: Open Large-scale Language Models
Aiyuan Yang, Bin Xiao, Bingning Wang +52
Large language models (LLMs) have demonstrated remarkable performance on a variety of natural language tasks based on just a few examples of natural language instructions, reducing…