3 papers
cs.DC2026
WISP: Waste- and Interference-Suppressed Distributed Speculative LLM Serving at the Edge via Dynamic Drafting and SLO-Aware Batching
Xiangchen Li, Jiakun Fan, Qingyuan Wang +7
As Large Language Models (LLMs) become increasingly accessible to end users, an ever-growing number of inference requests are initiated from edge devices and computed on centralize…
cs.AI2026
Intent-Driven Smart Manufacturing Integrating Knowledge Graphs and Large Language Models
Takoua Jradi, John Violos, Dimitrios Spatharakis +4
The increasing complexity of smart manufacturing environments demands interfaces that can translate high-level human intents into machine-executable actions. This paper presents a…
cs.DC2025
SLED: A Speculative LLM Decoding Framework for Efficient Edge Serving
Xiangchen Li, Dimitrios Spatharakis, Saeid Ghafouri +5
The growing gap between the increasing complexity of large language models (LLMs) and the limited computational budgets of edge devices poses a key challenge for efficient on-devic…