2 papers
cs.LG2025
Constrained Edge AI Deployment: Fine-Tuning vs Distillation for LLM Compression
Jacob Sander, David Moe, Achraf Cohen +3
Modern foundational models are often compressed via a combination of structured pruning and re-training to meet the strict compute, memory, and connectivity constraints of edge dep…
cs.LG2025
On Accelerating Edge AI: Optimizing Resource-Constrained Environments
Jacob Sander, Achraf Cohen, Venkat R. Dasari +2
Resource-constrained edge deployments demand AI solutions that balance high performance with stringent compute, memory, and energy limitations. In this survey, we present a compreh…