20 citations · 21 across the 3 of their papers we have counts for
3 papers
cs.CL2023★ 1 cited
SiRA: Sparse Mixture of Low Rank Adaptation
Yun Zhu, Nevan Wichers, Chu-Cheng Lin +8
Parameter Efficient Tuning has been an prominent approach to adapt the Large Language Model to downstream tasks. Most previous works considers adding the dense trainable parameters…
cs.CL2023
UT5: Pretraining Non autoregressive T5 with unrolled denoising
Mahmoud G. Salem, Jiayu Ye, Chu-Cheng Lin +1
Recent advances in Transformer-based Large Language Models have made great strides in natural language generation. However, to decode K tokens, an autoregressive model needs K sequ…
cs.NE2014★ 20 cited
Recurrent-Neural-Network for Language Detection on Twitter Code-Switching Corpus
Joseph Chee Chang, Chu-Cheng Lin
Mixed language data is one of the difficult yet less explored domains of natural language processing. Most research in fields like machine translation or sentiment analysis assume…