2 papers
cs.CL2025
Evaluating Long-Term Memory for Long-Context Question Answering
Alessandra Terranova, Björn Ross, Alexandra Birch
In order for large language models to achieve true conversational continuity and benefit from experiential learning, they need memory. While research has focused on the development…
cs.CL2025
Compositional Generalisation for Explainable Hate Speech Detection
Agostina Calabrese, Tom Sherborne, Björn Ross +1
Hate speech detection is key to online content moderation, but current models struggle to generalise beyond their training data. This has been linked to dataset biases and the use…