10 citations · 21 across the 19 of their papers we have counts for
4 papers · 1 filter
Erasing Without Remembering: Implicit Knowledge Forgetting in Large Language Models
Huazheng Wang, Yongcheng Jing, Haifeng Sun +4
In this paper, we investigate knowledge forgetting in large language models with a focus on its generalisation, ensuring that models forget not only specific training samples but a…
ChatTime: A Unified Multimodal Time Series Foundation Model Bridging Numerical and Textual Data
Chengsen Wang, Qi Qi, Jingyu Wang +5
Human experts typically integrate numerical and textual multimodal information to analyze time series. However, most traditional deep learning predictors rely solely on unimodal nu…
OutlierTune: Efficient Channel-Wise Quantization for Large Language Models
Jinguang Wang, Yuexi Yin, Haifeng Sun +5
Quantizing the activations of large language models (LLMs) has been a significant challenge due to the presence of structured outliers. Most existing methods focus on the per-token…
How Does Diffusion Influence Pretrained Language Models on Out-of-Distribution Data?
Huazheng Wang, Daixuan Cheng, Haifeng Sun +5
Transformer-based pretrained language models (PLMs) have achieved great success in modern NLP. An important advantage of PLMs is good out-of-distribution (OOD) robustness. Recently…