335 citations · 506 across the 13 of their papers we have counts for
6 papers · 1 filter
Inserting Information Bottlenecks for Attribution in Transformers
Zhiying Jiang, Raphael Tang, Ji Xin +1
Pretrained transformers achieve the state of the art across tasks in natural language processing, motivating researchers to investigate their inner mechanisms. One common direction…
Howl: A Deployed, Open-Source Wake Word Detection System
Raphael Tang, Jaejun Lee, Afsaneh Razi +4
We describe Howl, an open-source wake word detection toolkit with native support for open speech datasets, like Mozilla Common Voice and Google Speech Commands. We report benchmark…
Covidex: Neural Ranking Models and Keyword Search Infrastructure for the COVID-19 Open Research Dataset
Edwin Zhang, Nikhil Gupta, Raphael Tang +8
We present Covidex, a search engine that exploits the latest neural ranking models to provide information access to the COVID-19 Open Research Dataset curated by the Allen Institut…
Showing Your Work Doesn't Always Work
Raphael Tang, Jaejun Lee, Ji Xin +3
In natural language processing, a recently popular line of work explores how to best report the experimental results of neural networks. One exemplar publication, titled "Show Your…
DeeBERT: Dynamic Early Exiting for Accelerating BERT Inference
Ji Xin, Raphael Tang, Jaejun Lee +2
Large-scale pre-trained language models such as BERT have brought significant improvements to NLP applications. However, they are also notorious for being slow in inference, which…
Rapidly Bootstrapping a Question Answering Dataset for COVID-19
Raphael Tang, Rodrigo Nogueira, Edwin Zhang +4
We present CovidQA, the beginnings of a question answering dataset specifically designed for COVID-19, built by hand from knowledge gathered from Kaggle's COVID-19 Open Research Da…