activity
20182021
most citedVT-SSum: A Benchmark Dataset for Video Transcript Segmentation and Summarization

3 citations · 3 across the 1 of their papers we have counts for

collaborators
Showing cs.CLShow all

7 papers · 1 filter

cs.CL20213 cited

VT-SSum: A Benchmark Dataset for Video Transcript Segmentation and Summarization

Tengchao Lv, Lei Cui, Momcilo Vasilijevic +1

Video transcript summarization is a fundamental task for video understanding. Conventional approaches for transcript summarization are usually built upon the summarization data for…

cs.CL2020

DocBank: A Benchmark Dataset for Document Layout Analysis

Minghao Li, Yiheng Xu, Lei Cui +4

Document layout analysis usually relies on computer vision models to understand documents while ignoring textual information that is vital to capture. Meanwhile, high quality label…

cs.CL2019

LayoutLM: Pre-training of Text and Layout for Document Image Understanding

Yiheng Xu, Minghao Li, Lei Cui +3

Pre-training techniques have been verified successfully in a variety of NLP tasks in recent years. Despite the widespread use of pre-training models for NLP applications, they almo…

cs.CL2018

Unsupervised Machine Commenting with Neural Variational Topic Model

Shuming Ma, Lei Cui, Furu Wei +1

Article comments can provide supplementary opinions and facts for readers, thereby increase the attraction and engagement of articles. Therefore, automatically commenting is helpfu…

cs.CL2018

Neural Melody Composition from Lyrics

Hangbo Bao, Shaohan Huang, Furu Wei +5

In this paper, we study a novel task that learns to compose music from natural language. Given the lyrics as input, we propose a melody composition model that generates lyrics-cond…

cs.CL2018

LiveBot: Generating Live Video Comments Based on Visual and Textual Contexts

Shuming Ma, Lei Cui, Damai Dai +2

We introduce the task of automatic live commenting. Live commenting, which is also called `video barrage', is an emerging feature on online video sites that allows real-time commen…