3 papers
cs.LG2026
AMNESIA: A Large Scale Medical Unlearning Benchmark Suite with Disease-Informed Analysis
Saeedeh Davoudi, Reihaneh Iranmanesh, Ophir Frieder +1
Medical knowledge is continuously evolving. This creates a need to update or selectively forget information encoded in already-trained medical LLMs. Machine unlearning aims to remo…
cs.CL2026
TARAZ: Persian Short-Answer Question Benchmark for Cultural Evaluation of Language Models
Reihaneh Iranmanesh, Saeedeh Davoudi, Pasha Abrishamchian +2
This paper presents a comprehensive evaluation framework for assessing the cultural competence of large language models (LLMs) in Persian. Existing Persian cultural benchmarks rely…
cs.CL2024
Learning to Rank Salient Content for Query-focused Summarization
Sajad Sotudeh, Nazli Goharian
This study examines the potential of integrating Learning-to-Rank (LTR) with Query-focused Summarization (QFS) to enhance the summary relevance via content prioritization. Using a…