3 papers
cs.CR2026
SentinelRAG: Synthetic Sentinel Knowledge for RAG Database Copyright Protection
Tsun On Kwok, Xi Yang, Ki Sen Hung +2
Protecting proprietary RAG databases from unauthorized redistribution is challenging: existing watermarking methods either inject fabricated relations between real entities, pollut…
cs.CR2026
Into the Gray Zone: Domain Contexts Can Blur LLM Safety Boundaries
Ki Sen Hung, Xi Yang, Chang Liu +7
A central goal of LLM alignment is to balance helpfulness with harmlessness, yet these objectives conflict when the same knowledge serves both legitimate and malicious purposes. Th…
cs.HC2026
GrandGuard: Taxonomy, Benchmark, and Safeguards for Elderly-Chatbot Interaction Safety
Changxuan Fan, Xi Yang, Yueyuan Zheng +9
As older adults increasingly use LLM-based chatbots for companionship and assistance, a safety gap is emerging. Older adults may face vulnerabilities from social isolation, limited…