5 citations · 6 across the 3 of their papers we have counts for
3 papers
cs.CL2024
Enhancing Multiple Dimensions of Trustworthiness in LLMs via Sparse Activation Control
Yuxin Xiao, Chaoqun Wan, Yonggang Zhang +5
As the development and application of Large Language Models (LLMs) continue to advance rapidly, enhancing their trustworthiness and aligning them with human preferences has become…
cs.CL2024★ 1 cited
Interpreting and Improving Large Language Models in Arithmetic Calculation
Wei Zhang, Chaoqun Wan, Yonggang Zhang +4
Large language models (LLMs) have demonstrated remarkable potential across numerous applications and have shown an emergent ability to tackle complex reasoning tasks, such as mathe…
cs.LG2022★ 5 cited
Towards Lightweight Black-Box Attacks against Deep Neural Networks
Chenghao Sun, Yonggang Zhang, Wan Chaoqun +5
Black-box attacks can generate adversarial examples without accessing the parameters of target model, largely exacerbating the threats of deployed deep neural networks (DNNs). Howe…