2 papers
cs.LG2024
IceFormer: Accelerated Inference with Long-Sequence Transformers on CPUs
Yuzhen Mao, Martin Ester, Ke Li
One limitation of existing Transformer-based models is that they cannot handle very long sequences as input since their self-attention operations exhibit quadratic time and space c…
cs.LG2022
Augmenting Knowledge Transfer across Graphs
Yuzhen Mao, Jianhui Sun, Dawei Zhou
Given a resource-rich source graph and a resource-scarce target graph, how can we effectively transfer knowledge across graphs and ensure a good generalization performance? In many…