Showing cs.AIShow all
2 papers · 1 filter
cs.AI2025
SAKE: Steering Activations for Knowledge Editing
Marco Scialanga, Thibault Laugel, Vincent Grari +1
As Large Langue Models have been shown to memorize real-world facts, the need to update this knowledge in a controlled and efficient manner arises. Designed with these constraints…
cs.AI2025
Metric assessment protocol in the context of answer fluctuation on MCQ tasks
Ekaterina Goliakova, Xavier Renard, Marie-Jeanne Lesot +3
Using multiple-choice questions (MCQs) has become a standard for assessing LLM capabilities efficiently. A variety of metrics can be employed for this task. However, previous resea…