2 citations · 3 across the 2 of their papers we have counts for
2 papers
cs.LG2023★ 2 cited
Baselines for Identifying Watermarked Large Language Models
Leonard Tang, Gavin Uberti, Tom Shlomi
We consider the emerging problem of identifying the presence and use of watermarking schemes in widely used, publicly hosted, closed source large language models (LLMs). We introdu…
cs.LG2023★ 1 cited
Learning the Wrong Lessons: Inserting Trojans During Knowledge Distillation
Leonard Tang, Tom Shlomi, Alexander Cai
In recent years, knowledge distillation has become a cornerstone of efficiently deployed machine learning, with labs and industries using knowledge distillation to train models tha…