331 citations · 335 across the 6 of their papers we have counts for
Showing 2022Show all
2 papers · 1 filter
cs.CL2022★ 1 cited
When a sentence does not introduce a discourse entity, Transformer-based models still sometimes refer to it
Sebastian Schuster, Tal Linzen
Understanding longer narratives or participating in conversations requires tracking of discourse entities that have been mentioned. Indefinite noun phrases (NPs), such as 'a dog',…
cs.CL2022★ 1 cited
Coloring the Blank Slate: Pre-training Imparts a Hierarchical Inductive Bias to Sequence-to-sequence Models
Aaron Mueller, Robert Frank, Tal Linzen +2
Relations between words are governed by hierarchical structure rather than linear ordering. Sequence-to-sequence (seq2seq) models, despite their success in downstream NLP applicati…