6 citations · 30 across the 18 of their papers we have counts for
26 papers
A Benchmark and Dataset for Post-OCR text correction in Sanskrit
Ayush Maheshwari, Nikhil Singh, Amrith Krishna +1
Sanskrit is a classical language with about 30 million extant manuscripts fit for digitisation, available in written, printed or scannedimage forms. However, it is still considered…
Speeding up NAS with Adaptive Subset Selection
Vishak Prasad C, Colin White, Paarth Jain +2
A majority of recent developments in neural architecture search (NAS) have been aimed at decreasing the computational cost of various techniques without affecting their final perfo…
Partitioned Gradient Matching-based Data Subset Selection for Compute-Efficient Robust ASR Training
Ashish Mittal, Durga Sivasubramanian, Rishabh Iyer +2
Training state-of-the-art ASR systems such as RNN-T often has a high associated financial and environmental cost. Training with a subset of training data could mitigate this proble…
AutoML for Climate Change: A Call to Action
Renbo Tu, Nicholas Roberts, Vishak Prasad +7
The challenge that climate change poses to humanity has spurred a rapidly developing field of artificial intelligence research focused on climate change applications. The climate c…
DIAGNOSE: Avoiding Out-of-distribution Data using Submodular Information Measures
Suraj Kothawade, Akshit Srivastava, Venkat Iyer +2
Avoiding out-of-distribution (OOD) data is critical for training supervised machine learning models in the medical imaging domain. Furthermore, obtaining labeled medical data is di…
CLINICAL: Targeted Active Learning for Imbalanced Medical Image Classification
Suraj Kothawade, Atharv Savarkar, Venkat Iyer +3
Training deep learning models on medical datasets that perform well for all classes is a challenging task. It is often the case that a suboptimal performance is obtained on some cl…