5 citations · 5 across the 2 of their papers we have counts for
2 papers
cs.CL2023
Comparing Variation in Tokenizer Outputs Using a Series of Problematic and Challenging Biomedical Sentences
Christopher Meaney, Therese A Stukel, Peter C Austin +1
Background & Objective: Biomedical text data are increasingly available for research. Tokenization is an initial step in many biomedical text mining pipelines. Tokenization is the…
stat.ME2019★ 5 cited
Variance partitioning in multilevel models for count data
George Leckie, William Browne, Harvey Goldstein +2
A first step when fitting multilevel models to continuous responses is to explore the degree of clustering in the data. Researchers fit variance-component models and then report th…