1 citations · 1 across the 3 of their papers we have counts for
3 papers
Towards Understanding and Measuring COGNITIVE ATROPHY in LLM Behaviour
Abeer Badawi, Moyosoreoluwa Olatosi, Negin Baghbanzadeh +5
Recent incidents involving LLMs used for mental-health support reveal a critical evaluation gap: surface-level safety scores do not capture how models behave across realistic, emot…
MedSynth: Realistic, Synthetic Medical Dialogue-Note Pairs
Ahmad Rezaie Mianroodi, Amirali Rezaie, Niko Grisel Todorov +6
Physicians spend significant time documenting clinical encounters, a burden that contributes to professional burnout. To address this, robust automation tools for medical documenta…
Exploring the features used for summary evaluation by Human and GPT
Zahra Sadeghi, Evangelos Milios, Frank Rudzicz
Summary assessment involves evaluating how well a generated summary reflects the key ideas and meaning of the source text, requiring a deep understanding of the content. Large Lang…