22 citations · 23 across the 2 of their papers we have counts for
4 papers · 1 filter
PolyAlign: Conditional Human-Distribution Alignment
L. D. M. S. Sai Teja, Ufaq Khan, Sathira Silva +2
Post-training methods such as supervised fine-tuning (SFT) and preference optimization typically align language models toward a single global assistant behavior. While effective fo…
MultiWOZ -- A Large-Scale Multi-Domain Wizard-of-Oz Dataset for Task-Oriented Dialogue Modelling
Paweł Budzianowski, Tsung-Hsien Wen, Bo-Hsiang Tseng +4
Even though machine learning has become the major scene in dialogue research community, the real breakthrough has been blocked by the scale of data available. To address this funda…
Deep learning for language understanding of mental health concepts derived from Cognitive Behavioural Therapy
Lina Rojas-Barahona, Bo-Hsiang Tseng, Yinpei Dai +5
In recent years, we have seen deep learning and distributed representations of words and sentences make impact on a number of natural language processing tasks, such as similarity,…
Large-Scale Multi-Domain Belief Tracking with Knowledge Sharing
Osman Ramadan, Paweł Budzianowski, Milica Gašić
Robust dialogue belief tracking is a key component in maintaining good quality dialogue systems. The tasks that dialogue systems are trying to solve are becoming increasingly compl…