2 citations · 2 across the 2 of their papers we have counts for
2 papers
cs.CV2022
Augmenting Vision Language Pretraining by Learning Codebook with Visual Semantics
Xiaoyuan Guo, Jiali Duan, C. -C. Jay Kuo +2
Language modality within the vision language pretraining framework is innately discretized, endowing each word in the language vocabulary a semantic meaning. In contrast, visual mo…
eess.IV2021★ 2 cited
MedShift: identifying shift data for medical dataset curation
Xiaoyuan Guo, Judy Wawira Gichoya, Hari Trivedi +2
To curate a high-quality dataset, identifying data variance between the internal and external sources is a fundamental and crucial step. However, methods to detect shift or varianc…