2 papers
cs.LG2026
How Implicit Bias Accumulates and Propagates in LLM Long-term Memory
Yiming Ma, Lixu Wang, Lionel Z. Wang +6
Long-term memory mechanisms enable Large Language Models (LLMs) to maintain continuity and personalization across extended interaction lifecycles, but they also introduce new and u…
cs.CL2025
JiraiBench: A Bilingual Benchmark for Evaluating Large Language Models' Detection of Human Self-Destructive Behavior Content in Jirai Community
Yunze Xiao, Tingyu He, Lionel Z. Wang +6
This paper introduces JiraiBench, the first bilingual benchmark for evaluating large language models' effectiveness in detecting self-destructive content across Chinese and Japanes…