Showing cs.AIShow all
2 papers · 1 filter
cs.AI2025
NoisyNN: Exploring the Impact of Information Entropy Change in Learning Systems
Xiaowei Yu, Zhe Huang, Minheng Chen +3
We investigate the impact of entropy change in deep learning systems by noise injection at different levels, including the embedding space and the image. The series of models that…
cs.AI2025
Alignment and Safety in Large Language Models: Safety Mechanisms, Training Paradigms, and Emerging Challenges
Haoran Lu, Luyang Fang, Ruidong Zhang +47
Due to the remarkable capabilities and growing impact of large language models (LLMs), they have been deeply integrated into many aspects of society. Thus, ensuring their alignment…