4 papers
Challenges in Explaining Pretrained Clinical Text Classifiers
Kristian Miok, Matej Klemen, Blaz Å krlj +1
Explaining the predictions of neural models in clinical NLP remains a significant challenge, especially for complex tasks involving long, unstructured medical texts. While post-hoc…
Evaluating Metalinguistic Knowledge in Large Language Models across the World's Languages
TjaÅ¡a ArÄon, Matej Klemen, Marko Robnik-Å ikonja +1
LLMs are routinely evaluated on language use, yet their explicit knowledge about linguistic structure remains poorly understood. Existing linguistic benchmarks focus on narrow phen…
Towards Corpus-Grounded Agentic LLMs for Multilingual Grammatical Analysis
Matej Klemen, TjaÅ¡a ArÄon, Luka TerÄon +2
Empirical grammar research has become increasingly data-driven, but the systematic analysis of annotated corpora still requires substantial methodological and technical effort. We…
Neural spell-checker: Beyond words with synthetic data generation
Matej Klemen, Martin BožiÄ, Å pela Arhar Holdt +1
Spell-checkers are valuable tools that enhance communication by identifying misspelled words in written texts. Recent improvements in deep learning, and in particular in large lang…