7 citations · 7 across the 1 of their papers we have counts for
2 papers
cs.DL2022★ 7 cited
A large dataset of software mentions in the biomedical literature
Ana-Maria Istrate, Donghui Li, Dario Taraborelli +3
We describe the CZ Software Mentions dataset, a new dataset of software mentions in biomedical papers. Plain-text software mentions are extracted with a trained SciBERT model from…
cs.CL2022
A Distant Supervision Corpus for Extracting Biomedical Relationships Between Chemicals, Diseases and Genes
Dongxu Zhang, Sunil Mohan, Michaela Torkar +1
We introduce ChemDisGene, a new dataset for training and evaluating multi-class multi-label document-level biomedical relation extraction models. Our dataset contains 80k biomedica…