3 papers
cs.AI2026
Transferable Latency Prediction for Fast LLM Screening on Heterogeneous Edge Devices
Xiaolong Tu, Vinod K. Mishra, Venkat R. Dasari +2
Accurate latency prediction is critical for deploying large language models (LLMs) on heterogeneous edge devices, where inference latency is affected by model architecture, prompt…
cs.LG2025
Semantic Edge Computing and Semantic Communications in 6G Networks: A Unifying Survey and Research Challenges
Milin Zhang, Mohammad Abdi, Venkat R. Dasari +1
Semantic Edge Computing (SEC) and Semantic Communications (SemComs) have been proposed as viable approaches to achieve real-time edge-enabled intelligence in sixth-generation (6G)…
cs.LG2025
On Accelerating Edge AI: Optimizing Resource-Constrained Environments
Jacob Sander, Achraf Cohen, Venkat R. Dasari +2
Resource-constrained edge deployments demand AI solutions that balance high performance with stringent compute, memory, and energy limitations. In this survey, we present a compreh…