Showing cs.CLShow all
2 papers · 1 filter
cs.CL2024
What Evidence Do Language Models Find Convincing?
Alexander Wan, Eric Wallace, Dan Klein
Retrieval-augmented language models are being increasingly tasked with subjective, contentious, and conflicting queries such as "is aspartame linked to cancer". To resolve these am…
cs.CL2023★ 38 cited
Poisoning Language Models During Instruction Tuning
Alexander Wan, Eric Wallace, Sheng Shen +1
Instruction-tuned LMs such as ChatGPT, FLAN, and InstructGPT are finetuned on datasets that contain user-submitted examples, e.g., FLAN aggregates numerous open-source datasets and…