activity
20242026
most citedCurating Stopwords in Marathi: A TF-IDF Approach for Improved Text Analysis and Information Retrieval

3 citations · 6 across the 8 of their papers we have counts for

collaborators
Showing 2024 · cs.CLShow all

6 papers · 2 filters

cs.CL2024

Long Range Named Entity Recognition for Marathi Documents

Pranita Deshmukh, Nikita Kulkarni, Sanhita Kulkarni +3

The demand for sophisticated natural language processing (NLP) methods, particularly Named Entity Recognition (NER), has increased due to the exponential growth of Marathi-language…

cs.CL2024★ 1 cited

L3Cube-MahaSum: A Comprehensive Dataset and BART Models for Abstractive Text Summarization in Marathi

Pranita Deshmukh, Nikita Kulkarni, Sanhita Kulkarni +2

We present the MahaSUM dataset, a large-scale collection of diverse news articles in Marathi, designed to facilitate the training and evaluation of models for abstractive summariza…

cs.CL2024★ 1 cited

A Data Selection Approach for Enhancing Low Resource Machine Translation Using Cross-Lingual Sentence Representations

Nidhi Kowtal, Tejas Deshpande, Raviraj Joshi

Machine translation in low-resource language pairs faces significant challenges due to the scarcity of parallel corpora and linguistic resources. This study focuses on the case of…

cs.CL2024

Chain-of-Translation Prompting (CoTR): A Novel Prompting Technique for Low Resource Languages

Tejas Deshpande, Nidhi Kowtal, Raviraj Joshi

This paper introduces Chain of Translation Prompting (CoTR), a novel strategy designed to enhance the performance of language models in low-resource languages. CoTR restructures pr…

cs.CL2024★ 1 cited

Leveraging Parameter Efficient Training Methods for Low Resource Text Classification: A Case Study in Marathi

Pranita Deshmukh, Nikita Kulkarni, Sanhita Kulkarni +2

With the surge in digital content in low-resource languages, there is an escalating demand for advanced Natural Language Processing (NLP) techniques tailored to these languages. BE…

cs.CL2024★ 3 cited

Curating Stopwords in Marathi: A TF-IDF Approach for Improved Text Analysis and Information Retrieval

Rohan Chavan, Gaurav Patil, Vishal Madle +1

Stopwords are commonly used words in a language that are often considered to be of little value in determining the meaning or significance of a document. These words occur frequent…