most citedTraining Data Leakage Analysis in Language Models

22 citations · 23 across the 2 of their papers we have counts for

collaborators

5 papers

cs.CR202122 cited

Training Data Leakage Analysis in Language Models

Huseyin A. Inan, Osman Ramadan, Lukas Wutschitz +4

Recent advances in neural network based language models lead to successful deployments of such models, improving user experience in various applications. It has been demonstrated t…

cs.CR20211 cited

N-grams Bayesian Differential Privacy

Osman Ramadan, James Withers, Douglas Orr

Differential privacy has gained popularity in machine learning as a strong privacy guarantee, in contrast to privacy mitigation techniques such as k-anonymity. However, applying di…

cs.CL2018

MultiWOZ -- A Large-Scale Multi-Domain Wizard-of-Oz Dataset for Task-Oriented Dialogue Modelling

Paweł Budzianowski, Tsung-Hsien Wen, Bo-Hsiang Tseng +4

Even though machine learning has become the major scene in dialogue research community, the real breakthrough has been blocked by the scale of data available. To address this funda…

cs.CL2018

Deep learning for language understanding of mental health concepts derived from Cognitive Behavioural Therapy

Lina Rojas-Barahona, Bo-Hsiang Tseng, Yinpei Dai +5

In recent years, we have seen deep learning and distributed representations of words and sentences make impact on a number of natural language processing tasks, such as similarity,…

cs.CL2018

Large-Scale Multi-Domain Belief Tracking with Knowledge Sharing

Osman Ramadan, Paweł Budzianowski, Milica Gašić

Robust dialogue belief tracking is a key component in maintaining good quality dialogue systems. The tasks that dialogue systems are trying to solve are becoming increasingly compl…