4 papers
How to model Human Actions distribution with Event Sequence Data
Egor Surkov, Dmitry Osin, Evgeny Burnaev +1
This paper studies forecasting of the future distribution of events in human action sequences, a task essential in domains like retail, finance, healthcare, and recommendation syst…
Token Homogenization under Positional Bias
Viacheslav Yusupov, Danil Maksimov, Ameliia Alaeva +9
This paper investigates token homogenization - the convergence of token representations toward uniformity across transformer layers and its relationship to positional bias in large…
Beyond Early-Token Bias: Model-Specific and Language-Specific Position Effects in Multilingual LLMs
Mikhail Menschikov, Alexander Kharitonov, Maiia Kotyga +5
Large Language Models (LLMs) exhibit position bias systematically underweighting information based on its location in the context but how this bias varies across languages and mode…
Investigating the Impact of Quantization Methods on the Safety and Reliability of Large Language Models
Artyom Kharinaev, Viktor Moskvoretskii, Egor Shvetsov +3
Large Language Models (LLMs) are powerful tools for modern applications, but their computational demands limit accessibility. Quantization offers efficiency gains, yet its impact o…