10 citations · 26 across the 14 of their papers we have counts for
4 papers · 2 filters
Document-aligned Japanese-English Conversation Parallel Corpus
Matīss Rikters, Ryokan Ri, Tong Li +1
Sentence-level (SL) machine translation (MT) has reached acceptable quality for many high-resourced languages, but not document-level (DL) MT, which is difficult to 1) train with l…
Designing the Business Conversation Corpus
Matīss Rikters, Ryokan Ri, Tong Li +1
While the progress of machine translation of written text has come far in the past several years thanks to the increasing availability of parallel corpora and corpora-based trainin…
Data Augmentation with Unsupervised Machine Translation Improves the Structural Similarity of Cross-lingual Word Embeddings
Sosuke Nishikawa, Ryokan Ri, Yoshimasa Tsuruoka
Unsupervised cross-lingual word embedding (CLWE) methods learn a linear transformation matrix that maps two monolingual embedding spaces that are separately trained with monolingua…
Revisiting the Context Window for Cross-lingual Word Embeddings
Ryokan Ri, Yoshimasa Tsuruoka
Existing approaches to mapping-based cross-lingual word embeddings are based on the assumption that the source and target embedding spaces are structurally similar. The structures…