4 citations · 8 across the 2 of their papers we have counts for
3 papers
cs.CL2024
Learn To be Efficient: Build Structured Sparsity in Large Language Models
Haizhong Zheng, Xiaoyan Bai, Xueshen Liu +4
Large Language Models (LLMs) have achieved remarkable success with their billion-level parameters, yet they incur high inference overheads. The emergence of activation sparsity in…
cs.CL2024★ 4 cited
A Mechanistic Understanding of Alignment Algorithms: A Case Study on DPO and Toxicity
Andrew Lee, Xiaoyan Bai, Itamar Pres +3
While alignment algorithms are now commonly used to tune pre-trained language models towards a user's preferences, we lack explanations for the underlying mechanisms in which model…
cs.IR2023★ 4 cited
PromptRank: Unsupervised Keyphrase Extraction Using Prompt
Aobo Kong, Shiwan Zhao, Hao Chen +4
The keyphrase extraction task refers to the automatic selection of phrases from a given document to summarize its core content. State-of-the-art (SOTA) performance has recently bee…