activity
20182021
collaborators

7 papers

cs.CL2021

Language Identification of Hindi-English tweets using code-mixed BERT

Mohd Zeeshan Ansari, M M Sufyan Beg, Tanvir Ahmad +2

Language identification of social media text has been an interesting problem of study in recent years. Social media messages are predominantly in code mixed in non-English speaking…

cs.CL2021

Language Lexicons for Hindi-English Multilingual Text Processing

Mohd Zeeshan Ansari, Tanvir Ahmad, Noaima Bari

Language Identification in textual documents is the process of automatically detecting the language contained in a document based on its content. The present Language Identificatio…

cs.CL2021

A Simple and Efficient Probabilistic Language model for Code-Mixed Text

M Zeeshan Ansari, Tanvir Ahmad, M M Sufyan Beg +1

The conventional natural language processing approaches are not accustomed to the social media text due to colloquial discourse and non-homogeneous characteristics. Significantly,…

cs.SI2020

Inferring Political Preferences from Twitter

Mohd Zeeshan Ansari, Areesha Fatima Siddiqui, Mohammad Anas

Sentiment analysis is the task of automatic analysis of opinions and emotions of users towards an entity or some aspect of that entity. Political Sentiment Analysis of social media…

cs.CL2020

Feature Selection on Noisy Twitter Short Text Messages for Language Identification

Mohd Zeeshan Ansari, Tanvir Ahmad, Ana Fatima

The task of written language identification involves typically the detection of the languages present in a sample of text. Moreover, a sequence of text may not belong to a single i…

cs.CL2019

Context based Analysis of Lexical Semantics for Hindi Language

Mohd Zeeshan Ansari, Lubna Khan

A word having multiple senses in a text introduces the lexical semantic task to find out which particular sense is appropriate for the given context. One such task is Word sense di…