5 papers · 1 filter
Self-attention vector output similarities reveal how machines pay attention
Tal Halevi, Yarden Tzach, Ronit D. Gross +2
The self-attention mechanism has significantly advanced the field of natural language processing, facilitating the development of advanced language-learning machines. Although its…
Single-Nodal Spontaneous Symmetry Breaking in NLP Models
Shalom Rosner, Ronit D. Gross, Ella Koresh +1
Spontaneous symmetry breaking in statistical mechanics primarily occurs during phase transitions at the thermodynamic limit where the Hamiltonian preserves inversion symmetry, yet…
Translation Entropy: A Statistical Framework for Evaluating Translation Systems
Ronit D. Gross, Yanir Harel, Ido Kanter
The translation of written language has been known since the 3rd century BC; however, its necessity has become increasingly common in the information age. Today, many translators e…
Learning Mechanism Underlying NLP Pre-Training and Fine-Tuning
Yarden Tzach, Ronit D. Gross, Ella Koresh +4
Natural language processing (NLP) enables the understanding and generation of meaningful human language, typically using a pre-trained complex architecture on a large dataset to le…
Tiny language models
Ronit D. Gross, Yarden Tzach, Tal Halevi +2
A prominent achievement of natural language processing (NLP) is its ability to understand and generate meaningful human language. This capability relies on complex feedforward tran…