Showing 2024Show all
2 papers · 1 filter
cs.CL2024
In-Context Former: Lightning-fast Compressing Context for Large Language Model
Xiangfeng Wang, Zaiyi Chen, Zheyong Xie +3
With the rising popularity of Transformer-based large language models (LLMs), reducing their high inference costs has become a significant research focus. One effective approach is…
cs.LG2024
Boosting Gradient Ascent for Continuous DR-submodular Maximization
Qixin Zhang, Zongqi Wan, Zengde Deng +4
Projected Gradient Ascent (PGA) is the most commonly used optimization scheme in machine learning and operations research areas. Nevertheless, numerous studies and examples have sh…