Showing cs.CLShow all
2 papers · 1 filter
cs.CL2026
Community-Specific Slang and Entity Detection via Semantic Shift in Fine-Tuned Language Models
Julia Kruk, Sanchita Porwal, Amitrajit Bhattacharjee +1
We propose an unsupervised method of resolving slang, unique entities, and folklore from online communities by isolating words in the lexicon that have the highest magnitude of sem…
cs.CL2024
LLM Self Defense: By Self Examination, LLMs Know They Are Being Tricked
Mansi Phute, Alec Helbling, Matthew Hull +4
Large language models (LLMs) are popular for high-quality text generation but can produce harmful content, even when aligned with human values through reinforcement learning. Adver…