collaborators

7 papers

cs.CL2026

Pingala: Prosody-Aware Decoding for Sanskrit Poetry Generation

Manoj Balaji Jagadeeshan, Atul Singh, Nallani Chakravartula Sahith +2

Poetry generation in Sanskrit typically requires the verse to be semantically coherent and adhere to strict prosodic rules. In Sanskrit prosody, every line of a verse is typically…

cs.CL2026

Chandomitra: Towards Generating Structured Sanskrit Poetry from Natural Language Inputs

Manoj Balaji Jagadeeshan, Samarth Bhatia, Pretam Ray +7

Text Generation has achieved remarkable performance using large language models. It has also been recently well-studied that these large language models are capable of creative gen…

cs.CL2026

Mitrasamgraha: A Comprehensive Classical Sanskrit Machine Translation Dataset

Sebastian Nehrdich, David Allport, Sven Sellmer +5

While machine translation is regarded as a "solved problem" for many high-resource languages, close analysis quickly reveals that this is not the case for content that shows challe…

cs.CL2025

Still Not There: Can LLMs Outperform Smaller Task-Specific Seq2Seq Models on the Poetry-to-Prose Conversion Task?

Kunal Kingkar Das, Manoj Balaji Jagadeeshan, Nallani Chakravartula Sahith +2

Large Language Models (LLMs) are increasingly treated as universal, general-purpose solutions across NLP tasks, particularly in English. But does this assumption hold for low-resou…

cs.CL2025

Mahānāma: A Unique Testbed for Literary Entity Discovery and Linking

Sujoy Sarkar, Gourav Sarkar, Manoj Balaji Jagadeeshan +3

High lexical variation, ambiguous references, and long-range dependencies make entity resolution in literary texts particularly challenging. We present Mahānāma, the first large-…

cs.CL2025

Vedavani: A Benchmark Corpus for ASR on Vedic Sanskrit Poetry

Sujeet Kumar, Pretam Ray, Abhinay Beerukuri +3

Sanskrit, an ancient language with a rich linguistic heritage, presents unique challenges for automatic speech recognition (ASR) due to its phonemic complexity and the phonetic tra…