Showing cs.LGShow all
2 papers · 1 filter
cs.LG2025
Associative Transformer
Yuwei Sun, Hideya Ochiai, Zhirong Wu +2
Emerging from the pairwise attention in conventional Transformers, there is a growing interest in sparse attention mechanisms that align more closely with localized, contextual lea…
cs.LG2024
NuTime: Numerically Multi-Scaled Embedding for Large-Scale Time-Series Pretraining
Chenguo Lin, Xumeng Wen, Wei Cao +4
Recent research on time-series self-supervised models shows great promise in learning semantic representations. However, it has been limited to small-scale datasets, e.g., thousand…