1 citations · 1 across the 1 of their papers we have counts for
1 paper
Simeng Sun, Kalpesh Krishna, Andrew Mattarella-Micke +1
Language models are generally trained on short, truncated input sequences, which limits their ability to use discourse-level information present in long-range context to improve th…