activity
20232026
most citedAdaptive Feature-based Low-Rank Compression of Large Language Models via Bayesian Optimization

2 citations · 4 across the 16 of their papers we have counts for

collaborators
Showing 2024 · cs.CLShow all

5 papers · 2 filters

cs.CL2024

Revealing and Mitigating the Local Pattern Shortcuts of Mamba

Wangjie You, Zecheng Tang, Juntao Li +2

Large language models (LLMs) have advanced significantly due to the attention mechanism, but their quadratic complexity and linear memory demands limit their performance on long-co…

cs.CL2024

L-CiteEval: Do Long-Context Models Truly Leverage Context for Responding?

Zecheng Tang, Keyan Zhou, Juntao Li +3

Long-context models (LCMs) have made remarkable strides in recent years, offering users great convenience for handling tasks that involve long context, such as document summarizati…

cs.CL2024

MemLong: Memory-Augmented Retrieval for Long Text Modeling

Weijie Liu, Zecheng Tang, Juntao Li +2

Recent advancements in Large Language Models (LLMs) have yielded remarkable success across diverse fields. However, handling long contexts remains a significant challenge for LLMs…

cs.CL2024

Demonstration Augmentation for Zero-shot In-context Learning

Yi Su, Yunpeng Tai, Yixin Ji +3

Large Language Models (LLMs) have demonstrated an impressive capability known as In-context Learning (ICL), which enables them to acquire knowledge from textual demonstrations with…

cs.CL2024★ 2 cited

Adaptive Feature-based Low-Rank Compression of Large Language Models via Bayesian Optimization

Yixin Ji, Yang Xiang, Juntao Li +6

In recent years, large language models (LLMs) have driven advances in natural language processing. Still, their growing scale has increased the computational burden, necessitating…