1 citations · 1 across the 1 of their papers we have counts for
2 papers
cs.CL2021★ 1 cited
Do Long-Range Language Models Actually Use Long-Range Context?
Simeng Sun, Kalpesh Krishna, Andrew Mattarella-Micke +1
Language models are generally trained on short, truncated input sequences, which limits their ability to use discourse-level information present in long-range context to improve th…
cs.CL2020
Exploring and Predicting Transferability across NLP Tasks
Tu Vu, Tong Wang, Tsendsuren Munkhdalai +5
Recent advances in NLP demonstrate the effectiveness of training large-scale language models and transferring them to downstream tasks. Can fine-tuning these models on tasks other…