2 papers
cs.CL2024
Learning to Predict Usage Options of Product Reviews with LLM-Generated Labels
Leo Kohlenberg, Leonard Horns, Frederic Sadrieh +7
Annotating large datasets can be challenging. However, crowd-sourcing is often expensive and can lack quality, especially for non-trivial tasks. We propose a method of using LLMs a…
cs.CL2024
NextLevelBERT: Masked Language Modeling with Higher-Level Representations for Long Documents
Tamara Czinczoll, Christoph Hönes, Maximilian Schall +1
While (large) language models have significantly improved over the last years, they still struggle to sensibly process long sequences found, e.g., in books, due to the quadratic sc…