1 citations · 1 across the 5 of their papers we have counts for
Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
RAPTOR: Role-Aware Private Training for Mixture-of-Experts
Duc Dm, Khai Le-Duc, Nguyen Do +18
Differentially private (DP) fine-tuning methods treat sparse Mixture-of-Experts (MoE) models as a single dense block, ignoring that shared layers see all data while experts only se…
cs.LG2026
Semantic Optimal Transport for Sparse Autoencoder Feature Matching and Circuit Compression
Tue M. Cao, Nguyen Do, My T. Thai
Sparse autoencoders (SAEs) have become a central tool for interpreting language models. However, two key SAE analyses that remain difficult to scale are (1) matching semantically s…