113 citations · 188 across the 21 of their papers we have counts for
3 papers · 1 filter
Causal Estimation of Tokenisation Bias
Pietro Lesci, Clara Meister, Thomas Hofmann +2
Modern language models are typically trained over subword sequences, but ultimately define probabilities over character-strings. Ideally, the choice of the tokeniser -- which maps…
Explicit Word Density Estimation for Language Modelling
Jovan Andonov, Octavian Ganea, Paulina Grnarova +2
Language Modelling has been a central part of Natural Language Processing for a very long time and in the past few years LSTM-based language models have been the go-to method for c…
A Language Model's Guide Through Latent Space
Dimitri von Rütte, Sotiris Anagnostidis, Gregor Bachmann +1
Concept guidance has emerged as a cheap and simple way to control the behavior of language models by probing their hidden representations for concept vectors and using them to pert…