4 citations · 4 across the 2 of their papers we have counts for
2 papers
cs.CL2022
AD-DROP: Attribution-Driven Dropout for Robust Language Model Fine-Tuning
Tao Yang, Jinghao Deng, Xiaojun Quan +2
Fine-tuning large pre-trained language models on downstream tasks is apt to suffer from overfitting when limited training data is available. While dropout proves to be an effective…
cs.CL2022★ 4 cited
Autoregressive Entity Generation for End-to-End Task-Oriented Dialog
Guanhuan Huang, Xiaojun Quan, Qifan Wang
Task-oriented dialog (TOD) systems often require interaction with an external knowledge base to retrieve necessary entity (e.g., restaurant) information to support the response gen…