16 citations · 22 across the 2 of their papers we have counts for
3 papers
cs.IR2021★ 16 cited
Linear-Time Self Attention with Codeword Histogram for Efficient Recommendation
Yongji Wu, Defu Lian, Neil Zhenqiang Gong +4
Self-attention has become increasingly popular in a variety of sequence modeling tasks from natural language processing to recommendation, due to its effectiveness. However, self-a…
cs.IR2021★ 6 cited
Rethinking Lifelong Sequential Recommendation with Incremental Multi-Interest Attention
Yongji Wu, Lu Yin, Defu Lian +4
Sequential recommendation plays an increasingly important role in many e-commerce services such as display advertisement and online shopping. With the rapid development of these se…
cs.LG2021
Do We Actually Need Dense Over-Parameterization? In-Time Over-Parameterization in Sparse Training
Shiwei Liu, Lu Yin, Decebal Constantin Mocanu +1
In this paper, we introduce a new perspective on training deep neural networks capable of state-of-the-art performance without the need for the expensive over-parameterization by p…