Showing cs.LGShow all
3 papers · 1 filter
cs.LG2026
LLM-AutoDP: Automatic Data Processing via LLM Agents for Model Fine-tuning
Wei Huang, Anda Cheng, Yinggui Wang +2
Large Language Models (LLMs) can be fine-tuned on domain-specific data to enhance their performance in specialized fields. However, such data often contains numerous low-quality sa…
cs.LG2024
Information Leakage from Embedding in Large Language Models
Zhipeng Wan, Anda Cheng, Yinggui Wang +1
The widespread adoption of large language models (LLMs) has raised concerns regarding data privacy. This study aims to investigate the potential for privacy invasion through input…
cs.LG2024★ 4 cited
A Fast, Performant, Secure Distributed Training Framework For Large Language Model
Wei Huang, Yinggui Wang, Anda Cheng +3
The distributed (federated) LLM is an important method for co-training the domain-specific LLM using siloed data. However, maliciously stealing model parameters and data from the s…