18 citations · 47 across the 4 of their papers we have counts for
6 papers
WT-MVSNet: Window-based Transformers for Multi-view Stereo
Jinli Liao, Yikang Ding, Yoli Shavit +5
Recently, Transformers were shown to enhance the performance of multi-view stereo by enabling long-range feature interaction. In this work, we propose Window-based Transformers (WT…
Recursive Refinement Network for Deformable Lung Registration between Exhale and Inhale CT Scans
Xinzi He, Jia Guo, Xuzhe Zhang +9
Unsupervised learning-based medical image registration approaches have witnessed rapid development in recent years. We propose to revisit a commonly ignored while simple and well-e…
PTNet: A High-Resolution Infant MRI Synthesizer Based on Transformer
Xuzhe Zhang, Xinzi He, Jia Guo +6
Magnetic resonance imaging (MRI) noninvasively provides critical information about how human brain structures develop across stages of life. Developmental scientists are particular…
LAMP: Label Augmented Multimodal Pretraining
Jia Guo, Chen Zhu, Yilun Zhao +4
Multi-modal representation learning by pretraining has become an increasing interest due to its easy-to-use and potential benefit for various Visual-and-Language~(V-L) tasks. Howev…
Reducing the Teacher-Student Gap via Spherical Knowledge Disitllation
Jia Guo, Minghao Chen, Yao Hu +3
Knowledge distillation aims at obtaining a compact and effective model by learning the mapping function from a much larger one. Due to the limited capacity of the student, the stud…
MusiCoder: A Universal Music-Acoustic Encoder Based on Transformers
Yilun Zhao, Jia Guo
Music annotation has always been one of the critical topics in the field of Music Information Retrieval (MIR). Traditional models use supervised learning for music annotation tasks…