most citedARWKV: Pretrain is not what we need, an RNN-Attention-Based Language Model Born from Transformer

1 citations · 1 across the 5 of their papers we have counts for

collaborators

5 papers