3 papers
cs.SI2026
Understanding the Self-Reflection Mechanisms of LLMs through Biased Attitude Associations
Jingshen Zhang, Bo Wang, Boci Yang +4
While the emergent self-reflection capabilities of Large Language Models (LLMs) offer a promising paradigm for autonomous bias mitigation, their internal mechanics remain unclear,…
cs.SI2026
Modeling Implicit Conflict Monitoring Mechanisms against Stereotypes in LLMs
Jingshen Zhang, Bo Wang, Yanlin Fu +4
In this paper, we study an emergent self-debiasing mechanisms against stereotypical content in Large Language Models (LLMs). Unlike traditional safety mechanisms that are primarily…
cs.AI2026
A Multi-Memory Segment System for Generating High-Quality Long-Term Memory Content in Agents
Gaoke Zhang, Bo Wang, Yunlong Ma +2
In the current field of agent memory, extensive explorations have been conducted in the area of memory retrieval, yet few studies have focused on exploring the memory content. Most…