3 citations · 4 across the 14 of their papers we have counts for
Showing cs.LGShow all
2 papers · 1 filter
cs.LG2025
On-Demand Multi-Task Sparsity for Efficient Large-Model Deployment on Edge Devices
Lianming Huang, Haibo Hu, Qiao Li +2
Sparsity is essential for deploying large models on resource constrained edge platforms. However, optimizing sparsity patterns for individual tasks in isolation ignores the signifi…
cs.LG2025
Lossless Compression of Large Language Model-Generated Text via Next-Token Prediction
Yu Mao, Holger Pirk, Chun Jason Xue
As large language models (LLMs) continue to be deployed and utilized across domains, the volume of LLM-generated data is growing rapidly. This trend highlights the increasing impor…