12 citations · 13 across the 6 of their papers we have counts for
Showing cs.CRShow all
3 papers · 1 filter
cs.CR2026
Steering LLM Viewpoints through Fabricated Evidence Injection
Xi Yang, Chang Liu, Zhenglin Huang +4
As chatbots increasingly influence daily decision-making, their potential to produce misleading responses poses substantial risks to users. This paper investigates a critical cogni…
cs.CR2026
SentinelRAG: Synthetic Sentinel Knowledge for RAG Database Copyright Protection
Tsun On Kwok, Xi Yang, Ki Sen Hung +2
Protecting proprietary RAG databases from unauthorized redistribution is challenging: existing watermarking methods either inject fabricated relations between real entities, pollut…
cs.CR2026
Into the Gray Zone: Domain Contexts Can Blur LLM Safety Boundaries
Ki Sen Hung, Xi Yang, Chang Liu +7
A central goal of LLM alignment is to balance helpfulness with harmlessness, yet these objectives conflict when the same knowledge serves both legitimate and malicious purposes. Th…