4 papers
Steering LLM Viewpoints through Fabricated Evidence Injection
Xi Yang, Chang Liu, Zhenglin Huang +4
As chatbots increasingly influence daily decision-making, their potential to produce misleading responses poses substantial risks to users. This paper investigates a critical cogni…
SentinelRAG: Synthetic Sentinel Knowledge for RAG Database Copyright Protection
Tsun On Kwok, Xi Yang, Ki Sen Hung +2
Protecting proprietary RAG databases from unauthorized redistribution is challenging: existing watermarking methods either inject fabricated relations between real entities, pollut…
Into the Gray Zone: Domain Contexts Can Blur LLM Safety Boundaries
Ki Sen Hung, Xi Yang, Chang Liu +7
A central goal of LLM alignment is to balance helpfulness with harmlessness, yet these objectives conflict when the same knowledge serves both legitimate and malicious purposes. Th…
Exploiting Vulnerabilities in Speech Translation Systems through Targeted Adversarial Attacks
Chang Liu, Haolin Wu, Xi Yang +7
As speech translation (ST) systems become increasingly prevalent, understanding their vulnerabilities is crucial for ensuring robust and reliable communication. However, limited wo…