2 papers
cs.CL2026
Token Distillation: Attention-aware Input Embeddings For New Tokens
Konstantin Dobler, Desmond Elliott, Gerard de Melo
Current language models rely on static vocabularies determined at pretraining time, which can lead to decreased performance and increased computational cost for domains underrepres…
cs.CL2026
Segmentation and Processing of German Court Decisions from Open Legal Data
Harshil Darji, Martin Heckelmann, Christina Kratsch +1
The availability of structured legal data is important for advancing Natural Language Processing (NLP) techniques for the German legal system. One of the most widely used datasets,…