24 citations · 76 across the 63 of their papers we have counts for
9 papers · 2 filters
Utilizing Wordnets for Cognate Detection among Indian Languages
Diptesh Kanojia, Kevin Patel, Pushpak Bhattacharyya +2
Automatic Cognate Detection (ACD) is a challenging task which has been utilized to help NLP applications like Machine Translation, Information Retrieval and Computational Phylogene…
"A Passage to India": Pre-trained Word Embeddings for Indian Languages
Kumar Saurav, Kumar Saunack, Diptesh Kanojia +1
Dense word vectors or 'word embeddings' which encode semantic properties of words, have now become integral to NLP tasks like Machine Translation (MT), Question Answering (QA), Wor…
Challenge Dataset of Cognates and False Friend Pairs from Indian Languages
Diptesh Kanojia, Pushpak Bhattacharyya, Malhar Kulkarni +1
Cognates are present in multiple variants of the same text across different languages (e.g., "hund" in German and "hound" in English language mean "dog"). They pose a challenge to…
Harnessing Cross-lingual Features to Improve Cognate Detection for Low-resource Languages
Diptesh Kanojia, Raj Dabre, Shubham Dewangan +3
Cognates are variants of the same lexical form across different languages; for example 'fonema' in Spanish and 'phoneme' in English are cognates, both of which mean 'a unit of soun…
Cognition-aware Cognate Detection
Diptesh Kanojia, Prashant Sharma, Sayali Ghodekar +3
Automatic detection of cognates helps downstream NLP tasks of Machine Translation, Cross-lingual Information Retrieval, Computational Phylogenetics and Cross-lingual Named Entity R…
Automated Evidence Collection for Fake News Detection
Mrinal Rawat, Diptesh Kanojia
Fake news, misinformation, and unverifiable facts on social media platforms propagate disharmony and affect society, especially when dealing with an epidemic like COVID-19. The tas…