7 papers
Language Identification of Hindi-English tweets using code-mixed BERT
Mohd Zeeshan Ansari, M M Sufyan Beg, Tanvir Ahmad +2
Language identification of social media text has been an interesting problem of study in recent years. Social media messages are predominantly in code mixed in non-English speaking…
Language Lexicons for Hindi-English Multilingual Text Processing
Mohd Zeeshan Ansari, Tanvir Ahmad, Noaima Bari
Language Identification in textual documents is the process of automatically detecting the language contained in a document based on its content. The present Language Identificatio…
A Simple and Efficient Probabilistic Language model for Code-Mixed Text
M Zeeshan Ansari, Tanvir Ahmad, M M Sufyan Beg +1
The conventional natural language processing approaches are not accustomed to the social media text due to colloquial discourse and non-homogeneous characteristics. Significantly,…
Inferring Political Preferences from Twitter
Mohd Zeeshan Ansari, Areesha Fatima Siddiqui, Mohammad Anas
Sentiment analysis is the task of automatic analysis of opinions and emotions of users towards an entity or some aspect of that entity. Political Sentiment Analysis of social media…
Feature Selection on Noisy Twitter Short Text Messages for Language Identification
Mohd Zeeshan Ansari, Tanvir Ahmad, Ana Fatima
The task of written language identification involves typically the detection of the languages present in a sample of text. Moreover, a sequence of text may not belong to a single i…
Context based Analysis of Lexical Semantics for Hindi Language
Mohd Zeeshan Ansari, Lubna Khan
A word having multiple senses in a text introduces the lexical semantic task to find out which particular sense is appropriate for the given context. One such task is Word sense di…