activity
20242026
most citedLongHealth: A Question Answering Benchmark with Long Clinical Documents

2 citations · 3 across the 6 of their papers we have counts for

collaborators
Showing cs.CLShow all

5 papers · 1 filter

cs.CL20261 cited

A Sovereign, Open-Source Foundation Model for German and English

Soofi-Team, :, Benedikt Droste +30

We present Soofi S 30B-A3B, a sovereign, open-source Mixture-of-Experts (MoE) hybrid Mamba Transformer foundation model for German and English. Its hybrid design activates only 3B…

cs.CL2026

ReasonXL: Shifting LLM Reasoning Language Without Sacrificing Performance

Daniil Gurgurov, Tom Röhr, Sebastian von Rohrscheidt +3

Despite advances in multilingual capabilities, most large language models (LLMs) remain English-centric in their training and, crucially, in their production of reasoning traces. E…

cs.CL2026

Same Meaning, Different Scores: Lexical and Syntactic Sensitivity in LLM Evaluation

Bogdan Kostić, Conor Fallon, Julian Risch +1

The rapid advancement of Large Language Models (LLMs) has established standardized evaluation benchmarks as the primary instrument for model comparison. Yet, their reliability is i…

cs.CL2025

Comply: Learning Sentences with Complex Weights inspired by Fruit Fly Olfaction

Alexei Figueroa, Justus Westerhoff, Golzar Atefi +5

Biologically inspired neural networks offer alternative avenues to model data distributions. FlyVec is a recent example that draws inspiration from the fruit fly's olfactory circui…

cs.CL20242 cited

LongHealth: A Question Answering Benchmark with Long Clinical Documents

Lisa Adams, Felix Busch, Tianyu Han +7

Background: Recent advancements in large language models (LLMs) offer potential benefits in healthcare, particularly in processing extensive patient records. However, existing benc…