3 papers
cs.AI2026
What Can We Actually Steer? A Multi-Behavior Study of Activation Control
Tetiana Bas, Krystian Novak
Large language models (LLMs) require precise behavior control for safe and effective deployment across diverse applications. Activation steering offers a promising approach for LLM…
cs.CL2024
Benchmarking Multimodal Models for Ukrainian Language Understanding Across Academic and Cultural Domains
Yurii Paniv, Artur Kiulian, Dmytro Chaplynskyi +4
While the evaluation of multimodal English-centric models is an active area of research with numerous benchmarks, there is a profound lack of benchmarks or evaluation suites for lo…
cs.CL2024
Assessing Gender Bias in LLMs: Comparing LLM Outputs with Human Perceptions and Official Statistics
Tetiana Bas
This study investigates gender bias in large language models (LLMs) by comparing their gender perception to that of human respondents, U.S. Bureau of Labor Statistics data, and a 5…