2 papers
cs.CL2026
Can Continual Pre-training Bridge the Performance Gap between General-purpose and Specialized Language Models in the Medical Domain?
Niclas Doll, Jasper Schulze Buschhoff, Shalaka Satheesh +3
This paper narrows the performance gap between small, specialized models and significantly larger general-purpose models through domain adaptation via continual pre-training and me…
cs.CL2025
GG-BBQ: German Gender Bias Benchmark for Question Answering
Shalaka Satheesh, Katrin Klug, Katharina Beckh +3
Within the context of Natural Language Processing (NLP), fairness evaluation is often associated with the assessment of bias and reduction of associated harm. In this regard, the e…