3 citations · 5 across the 2 of their papers we have counts for
2 papers
cs.LG2023★ 2 cited
Mini-Batch Optimization of Contrastive Loss
Jaewoong Cho, Kartik Sreenivasan, Keon Lee +7
Contrastive learning has gained significant attention as a method for self-supervised learning. The contrastive loss function ensures that embeddings of positive sample pairs (e.g.…
cs.LG2023★ 3 cited
Looped Transformers as Programmable Computers
Angeliki Giannou, Shashank Rajput, Jy-yong Sohn +3
We present a framework for using transformer networks as universal computers by programming them with specific weights and placing them in a loop. Our input sequence acts as a punc…