4 citations · 5 across the 6 of their papers we have counts for
Showing cs.CLShow all
2 papers · 1 filter
cs.CL2023★ 4 cited
LoFT: Local Proxy Fine-tuning For Improving Transferability Of Adversarial Attacks Against Large Language Model
Muhammad Ahmed Shah, Roshan Sharma, Hira Dhamyal +10
It has been shown that Large Language Model (LLM) alignments can be circumvented by appending specially crafted attack suffixes with harmful queries to elicit harmful responses. To…
cs.CL2022★ 1 cited
Adapting Task-Oriented Dialogue Models for Email Conversations
Soham Deshmukh, Charles Lee
Intent detection is a key part of any Natural Language Understanding (NLU) system of a conversational assistant. Detecting the correct intent is essential yet difficult for email c…