1 citations · 1 across the 3 of their papers we have counts for
3 papers
cs.LG2023★ 1 cited
r-softmax: Generalized Softmax with Controllable Sparsity Rate
Klaudia Bałazy, Łukasz Struski, Marek Śmieja +1
Nowadays artificial neural network models achieve remarkable results in many disciplines. Functions mapping the representation provided by the model to the probability distribution…
cs.CL2023
Step by Step Loss Goes Very Far: Multi-Step Quantization for Adversarial Text Attacks
Piotr Gaiński, Klaudia Bałazy
We propose a novel gradient-based attack against transformer-based language models that searches for an adversarial example in a continuous space of token probabilities. Our algori…
cs.CL2023
Revisiting Offline Compression: Going Beyond Factorization-based Methods for Transformer Language Models
Mohammadreza Banaei, Klaudia Bałazy, Artur Kasymov +3
Recent transformer language models achieve outstanding results in many natural language processing (NLP) tasks. However, their enormous size often makes them impractical on memory-…