190 citations · 226 across the 7 of their papers we have counts for
3 papers · 1 filter
FlexMoE: One-for-All Nested Intra-Expert Pruning for MoE Language Models
Fan Mo, Yuxuan Han, Geng Zhang +2
Mixture-of-Experts (MoE) language models scale model ability with sparsely activated experts, making this architecture a standard recipe for modern large models. However, sparse ac…
Towards Battery-Free Machine Learning and Inference in Underwater Environments
Yuchen Zhao, Sayed Saad Afzal, Waleed Akbar +5
This paper is motivated by a simple question: Can we design and build battery-free devices capable of machine learning and inference in underwater environments? An affirmative answ…
DarkneTZ: Towards Model Privacy at the Edge using Trusted Execution Environments
Fan Mo, Ali Shahin Shamsabadi, Kleomenis Katevas +4
We present DarkneTZ, a framework that uses an edge device's Trusted Execution Environment (TEE) in conjunction with model partitioning to limit the attack surface against Deep Neur…