2 papers
cs.AI2026
Multi-Objective Structured Pruning of LLMs for Latency and Model Size Optimization
Muhammad Junaid Ali, Smail Niar, El-Ghazali Talbi
Large Language Models (LLMs) have achieved widespread adoption because of their strong reasoning and query-response capabilities. However, deploying them in embedded and edge compu…
cs.LG2024
Combining Neural Architecture Search and Automatic Code Optimization: A Survey
Inas Bachiri, Hadjer Benmeziane, Smail Niar +3
Deep Learning models have experienced exponential growth in complexity and resource demands in recent years. Accelerating these models for efficient execution on resource-constrain…