2 papers
cs.AI2026
Navigating the Sea of LLM Evaluation: Investigating Bias in Toxicity Benchmarks
Regina Gugg, Selina Niederländer, Andreas Stöckl +1
The rapid adoption of LLMs in both research and industry highlights the challenges of deploying them safely and reveals a gap in the systematic evaluation of toxicity benchmarks. A…
cs.CL2026
Linear-Time and Constant-Memory Text Embeddings Based on Recurrent Language Models
Tobias Grantner, Emanuel Sallinger, Martin Flechl
Transformer-based embedding models suffer from quadratic computational and linear memory complexity, limiting their utility for long sequences. We propose recurrent architectures a…