6 citations · 6 across the 1 of their papers we have counts for
3 papers
cs.LG2021★ 6 cited
Which transformer architecture fits my data? A vocabulary bottleneck in self-attention
Noam Wies, Yoav Levine, Daniel Jannai +1
After their successful debut in natural language processing, Transformer architectures are now becoming the de-facto standard in many domains. An obstacle for their deployment over…
cs.LG2020
The Depth-to-Width Interplay in Self-Attention
Yoav Levine, Noam Wies, Or Sharir +2
Self-attention architectures, which are rapidly pushing the frontier in natural language processing, demonstrate a surprising depth-inefficient behavior: previous works indicate th…
cond-mat.dis-nn2019
Deep autoregressive models for the efficient variational simulation of many-body quantum systems
Or Sharir, Yoav Levine, Noam Wies +2
Artificial Neural Networks were recently shown to be an efficient representation of highly-entangled many-body quantum states. In practical applications, neural-network states inhe…