2 papers
cs.CL2020
Unigram-Normalized Perplexity as a Language Model Performance Measure with Different Vocabulary Sizes
Jihyeon Roh, Sang-Hoon Oh, Soo-Young Lee
Although Perplexity is a widely used performance metric for language models, the values are highly dependent upon the number of words in the corpus and is useful to compare perform…
cs.CL2020
Hierarchical GPT with Congruent Transformers for Multi-Sentence Language Models
Jihyeon Roh, Huiseong Gim, Soo-Young Lee
We report a GPT-based multi-sentence language model for dialogue generation and document understanding. First, we propose a hierarchical GPT which consists of three blocks, i.e., a…