4 citations · 6 across the 6 of their papers we have counts for
Showing 2018Show all
2 papers · 1 filter
cs.CL2018
Unseen Word Representation by Aligning Heterogeneous Lexical Semantic Spaces
Victor Prokhorov, Mohammad Taher Pilehvar, Dimitri Kartsaklis +2
Word embedding techniques heavily rely on the abundance of training data for individual words. Given the Zipfian distribution of words in natural language texts, a large number of…
cs.CL2018
Card-660: Cambridge Rare Word Dataset - a Reliable Benchmark for Infrequent Word Representation Models
Mohammad Taher Pilehvar, Dimitri Kartsaklis, Victor Prokhorov +1
Rare word representation has recently enjoyed a surge of interest, owing to the crucial role that effective handling of infrequent words can play in accurate semantic understanding…