Showing cs.CLShow all
3 papers · 1 filter
cs.CL2026
Do Gender Cues Affect LLM Value Trade-offs? Evidence from a Controlled Decision Benchmark
Yangyang Liu, Dong Yu, Pengyuan Liu
Large language models are increasingly used in value-sensitive decision settings, where irrelevant demographic cues should not alter judgments. We construct the Realistic Value Dec…
cs.CL2026
THRD: A Training-Free Multi-Turn Defense Framework for Jailbreak Attacks on Large Language Models
Zhiqing Ma, Zhonghao Xu, Dong Yu +3
Multi-turn jailbreak attacks pose a growing threat to LLMs by exploiting conversational dynamics such as gradual escalation and cross-turn coordination. Existing defenses either re…
cs.CL2026
SPAGBias: Uncovering and Tracing Structured Spatial Gender Bias in Large Language Models
Binxian Su, Haoye Lou, Shucheng Zhu +4
Large language models (LLMs) are being increasingly used in urban planning, but since gendered space theory highlights how gender hierarchies are embedded in spatial organization,…