56 citations · 76 across the 14 of their papers we have counts for
1 paper · 1 filter
Dewen Zeng, Nan Du, Tao Wang +4
Overparameterized large-scale language models have impressive generalization performance of in-context few-shot learning. However, most language models allocate the same amount of…