1 paper
Shiyi Zhu, Jing Ye, Wei Jiang +4
Self-attention and position embedding are two key modules in transformer-based Large Language Models (LLMs). However, the potential relationship between them is far from well studi…