1 paper · 1 filter
Mathis Pink, Vy A. Vo, Qinyuan Wu +7
Current LLM benchmarks focus on evaluating models' memory of facts and semantic relations, primarily assessing semantic aspects of long-term memory. However, in humans, long-term m…