1 citations · 1 across the 2 of their papers we have counts for
5 papers
AutoLV: Automatic Lecture Video Generator
Wenbin Wang, Yang Song, Sanjay Jha
We propose an end-to-end lecture video generation system that can generate realistic and complete lecture videos directly from annotated slides, instructor's reference voice and in…
PhotoChat: A Human-Human Dialogue Dataset with Photo Sharing Behavior for Joint Image-Text Modeling
Xiaoxue Zang, Lijuan Liu, Maria Wang +3
We present a new human-human dialogue dataset - PhotoChat, the first dataset that casts light on the photo sharing behavior in onlin emessaging. PhotoChat contains 12k dialogues, e…
Fast WordPiece Tokenization
Xinying Song, Alex Salcianu, Yang Song +2
Tokenization is a fundamental preprocessing step for almost all NLP tasks. In this paper, we propose efficient algorithms for the WordPiece tokenization used in BERT, from single-w…
Extremely Small BERT Models from Mixed-Vocabulary Training
Sanqiang Zhao, Raghav Gupta, Yang Song +1
Pretrained language models like BERT have achieved good results on NLP tasks, but are impractical on resource-limited devices due to memory footprint. A large fraction of this foot…
Deep Dual Pyramid Network for Barcode Segmentation using Barcode-30k Database
Qijie Zhao, Feng Ni, Yang Song +2
Digital signs(such as barcode or QR code) are widely used in our daily life, and for many applications, we need to localize them on images. However, difficult cases such as targets…