2 citations · 2 across the 1 of their papers we have counts for
2 papers
cs.CL2024★ 2 cited
German Text Embedding Clustering Benchmark
Silvan Wehrli, Bert Arnrich, Christopher Irrgang
This work introduces a benchmark assessing the performance of clustering German text embeddings in different domains. This benchmark is driven by the increasing use of clustering n…
cs.CL2023
xMEN: A Modular Toolkit for Cross-Lingual Medical Entity Normalization
Florian Borchert, Ignacio Llorca, Roland Roller +2
Objective: To improve performance of medical entity normalization across many languages, especially when fewer language resources are available compared to English. Materials and M…