3 papers
cs.CL2025
The truth is no diaper: Human and AI-generated associations to emotional words
Špela Vintar, Jan Jona Javoršek
Human word associations are a well-known method of gaining insight into the internal mental lexicon, but the responses spontaneously offered by human participants to word cues are…
cs.CL2025
Charting the European LLM Benchmarking Landscape: A New Taxonomy and a Set of Best Practices
Špela Vintar, Taja Kuzman Pungeršek, Mojca Brglez +1
While new benchmarks for large language models (LLMs) are being developed continuously to catch up with the growing capabilities of new models and AI in general, using and evaluati…
cs.CL2025
From Polyester Girlfriends to Blind Mice: Creating the First Pragmatics Understanding Benchmarks for Slovene
Mojca Brglez, Å pela Vintar
Large language models are demonstrating increasing capabilities, excelling at benchmarks once considered very difficult. As their capabilities grow, there is a need for more challe…