activity
20162022
most citedRecent Advances in Natural Language Processing via Large Pre-Trained Language Models: A Survey

161 citations · 214 across the 21 of their papers we have counts for

collaborators

34 papers

cs.CL2022

MINION: a Large-Scale and Diverse Dataset for Multilingual Event Detection

Amir Pouran Ben Veyseh, Minh Van Nguyen, Franck Dernoncourt +1

Event Detection (ED) is the task of identifying and classifying trigger words of event mentions in text. Despite considerable research efforts in recent years for English text, the…

cs.CL2022

MEE: A Novel Multilingual Event Extraction Dataset

Amir Pouran Ben Veyseh, Javid Ebrahimi, Franck Dernoncourt +1

Event Extraction (EE) is one of the fundamental tasks in Information Extraction (IE) that aims to recognize event mentions and their arguments (i.e., participants) from text. Due t…

cs.CL202215 cited

Textual Data Augmentation for Patient Outcomes Prediction

Qiuhao Lu, Dejing Dou, Thien Huu Nguyen

Deep learning models have demonstrated superior performance in various healthcare applications. However, the major limitation of these deep models is usually the lack of high-quali…

cs.CL2022

Tutorial Recommendation for Livestream Videos using Discourse-Level Consistency and Ontology-Based Filtering

Amir Pouran Ben Veyseh, Franck Dernoncourt, Thien Huu Nguyen

Streaming videos is one of the methods for creators to share their creative works with their audience. In these videos, the streamer share how they achieve their final objective by…

cs.CL20221 cited

Improving Keyphrase Extraction with Data Augmentation and Information Filtering

Amir Pouran Ben Veyseh, Nicole Meister, Franck Dernoncourt +1

Keyphrase extraction is one of the essential tasks for document understanding in NLP. While the majority of the prior works are dedicated to the formal setting, e.g., books, news o…

cs.CL20222 cited

Symlink: A New Dataset for Scientific Symbol-Description Linking

Viet Dac Lai, Amir Pouran Ben Veyseh, Franck Dernoncourt +1

Mathematical symbols and descriptions appear in various forms across document section boundaries without explicit markup. In this paper, we present a new large-scale dataset that e…