activity
20162024
most citedOCR++: A Robust Framework For Information Extraction from Scholarly Articles

15 citations · 22 across the 10 of their papers we have counts for

collaborators
Showing cs.CLShow all

5 papers · 1 filter

cs.CL2024

How Robust are the Tabular QA Models for Scientific Tables? A Study using Customized Dataset

Akash Ghosh, B Venkata Sahith, Niloy Ganguly +2

Question-answering (QA) on hybrid scientific tabular and textual data deals with scientific information, and relies on complex numerical reasoning. In recent years, while tabular Q…

cs.CL2024

Cross-lingual Editing in Multilingual Language Models

Himanshu Beniwal, Kowsik Nandagopan D, Mayank Singh

The training of large language models (LLMs) necessitates substantial data and computational resources, and updating outdated LLMs entails significant efforts and resources. While…

cs.CL2023

Unveiling the Multi-Annotation Process: Examining the Influence of Annotation Quantity and Instance Difficulty on Model Performance

Pritam Kadasi, Mayank Singh

The NLP community has long advocated for the construction of multi-annotator datasets to better capture the nuances of language interpretation, subjectivity, and ambiguity. This pa…

cs.CL2023

Unlocking Model Insights: A Dataset for Automated Model Card Generation

Shruti Singh, Hitesh Lodwal, Husain Malwat +2

Language models (LMs) are no longer restricted to ML community, and instruction-tuned LMs have led to a rise in autonomous AI agents. As the accessibility of LMs grows, it is imper…

cs.CL2023

MUTANT: A Multi-sentential Code-mixed Hinglish Dataset

Rahul Gupta, Vivek Srivastava, Mayank Singh

The multi-sentential long sequence textual data unfolds several interesting research directions pertaining to natural language processing and generation. Though we observe several…