4 citations · 4 across the 1 of their papers we have counts for
3 papers
cs.LG2024
DART-Eval: A Comprehensive DNA Language Model Evaluation Benchmark on Regulatory DNA
Aman Patel, Arpita Singhal, Austin Wang +3
Recent advances in self-supervised models for natural language, vision, and protein sequences have inspired the development of large genomic DNA language models (DNALMs). These mod…
cs.LG2020★ 4 cited
Unsupervised Calibration under Covariate Shift
Anusri Pampari, Stefano Ermon
A probabilistic model is said to be calibrated if its predicted probabilities match the corresponding empirical frequencies. Calibration is important for uncertainty quantification…
cs.CL2018
emrQA: A Large Corpus for Question Answering on Electronic Medical Records
Anusri Pampari, Preethi Raghavan, Jennifer Liang +1
We propose a novel methodology to generate domain-specific large-scale question answering (QA) datasets by re-purposing existing annotations for other NLP tasks. We demonstrate an…