2 papers
cs.CL2026
Kwai Summary Attention Technical Report
Chenglong Chu, Guorui Zhou, Guowang Zhang +35
Long-context ability, has become one of the most important iteration direction of next-generation Large Language Models, particularly in semantic understanding/reasoning, code agen…
cs.LG2024
Graph Propagation Transformer for Graph Representation Learning
Zhe Chen, Hao Tan, Tao Wang +5
This paper presents a novel transformer architecture for graph representation learning. The core insight of our method is to fully consider the information propagation among nodes…