activity
20152026
most citedDistilling Task-Specific Knowledge from BERT into Simple Neural Networks

335 citations · 580 across the 57 of their papers we have counts for

collaborators
Showing 2021Show all

7 papers · 1 filter

cs.CL2021

Search and Learn: Improving Semantic Coverage for Data-to-Text Generation

Shailza Jolly, Zi Xuan Zhang, Andreas Dengel +1

Data-to-text generation systems aim to generate text descriptions based on input data (often represented in the tabular form). A typical system uses huge training samples for learn…

cs.CL2021★ 23 cited

Non-Autoregressive Translation with Layer-Wise Prediction and Deep Supervision

Chenyang Huang, Hao Zhou, Osmar R. Zaïane +2

How do we perform efficient inference while retaining high translation quality? Existing neural machine translation models, such as Transformer, achieve high performance, but they…

cs.LG2021★ 33 cited

Simulated annealing for optimization of graphs and sequences

Xianggen Liu, Pengyong Li, Fandong Meng +5

Optimization of discrete structures aims at generating a new structure with the better property given an existing one, which is a fundamental problem in machine learning. Different…

cs.CL2021★ 3 cited

Simulated Annealing for Emotional Dialogue Systems

Chengzhang Dong, Chenyang Huang, Osmar Zaïane +1

Explicitly modeling emotions in dialogue generation has important applications, such as building empathetic personal companions. In this study, we consider the task of expressing a…

cs.CL2021

Weakly Supervised Explainable Phrasal Reasoning with Neural Fuzzy Logic

Zijun Wu, Zi Xuan Zhang, Atharva Naik +3

Natural language inference (NLI) aims to determine the logical relationship between two sentences, such as Entailment, Contradiction, and Neutral. In recent years, deep learning mo…

cs.CL2021

A Globally Normalized Neural Model for Semantic Parsing

Chenyang Huang, Wei Yang, Yanshuai Cao +2

In this paper, we propose a globally normalized model for context-free grammar (CFG)-based semantic parsing. Instead of predicting a probability, our model predicts a real-valued s…