330 citations · 375 across the 42 of their papers we have counts for
6 papers · 1 filter
CLARITY: Clinical Assistant for Routing, Inference, and Triage
Vladimir Shaposhnikov, Aleksandr Nesterov, Ilia Kopanichuk +7
We present CLARITY (Clinical Assistant for Routing, Inference and Triage), an AI-driven platform designed to facilitate patient-to-specialist routing, clinical consultations, and s…
Geopolitical biases in LLMs: what are the "good" and the "bad" countries according to contemporary language models
Mikhail Salnikov, Dmitrii Korzh, Ivan Lazichny +7
This paper evaluates geopolitical biases in LLMs with respect to various countries though an analysis of their interpretation of historical events with conflicting national perspec…
Confidence Is All You Need: Few-Shot RL Fine-Tuning of Language Models
Pengyi Li, Matvey Skripkin, Alexander Zubrey +2
Large language models (LLMs) excel at reasoning, yet post-training remains critical for aligning their behavior with task goals. Existing reinforcement learning (RL) methods often…
On the Spatial Structure of Mixture-of-Experts in Transformers
Daniel Bershatsky, Ivan Oseledets
A common assumption is that MoE routers primarily leverage semantic features for expert selection. However, our study challenges this notion by demonstrating that positional token…
LoTR: Low Tensor Rank Weight Adaptation
Daniel Bershatsky, Daria Cherniuk, Talgat Daulbaev +2
In this paper we generalize and extend an idea of low-rank adaptation (LoRA) of large language models (LLMs) based on Transformer architecture. Widely used LoRA-like methods of fin…
Translate your gibberish: black-box adversarial attack on machine translation systems
Andrei Chertkov, Olga Tsymboi, Mikhail Pautov +1
Neural networks are deployed widely in natural language processing tasks on the industrial scale, and perhaps the most often they are used as compounds of automatic machine transla…