Showing cs.CLShow all
2 papers · 1 filter
cs.CL2024
Benchmarking Multimodal Models for Ukrainian Language Understanding Across Academic and Cultural Domains
Yurii Paniv, Artur Kiulian, Dmytro Chaplynskyi +4
While the evaluation of multimodal English-centric models is an active area of research with numerous benchmarks, there is a profound lack of benchmarks or evaluation suites for lo…
cs.CL2024
Assessing Gender Bias in LLMs: Comparing LLM Outputs with Human Perceptions and Official Statistics
Tetiana Bas
This study investigates gender bias in large language models (LLMs) by comparing their gender perception to that of human respondents, U.S. Bureau of Labor Statistics data, and a 5…