13 papers
An Internet for the KV Cache: Rethinking Classical Infrastructure Boundaries in the LLM Inference Age
Siddhant Ray, Nick Feamster, Junchen Jiang
LLM inference has become a global-scale, heterogeneous workload spanning agents, retrieval, tool-use, code execution and multi-modal reasoning. These workloads naturally enable con…
Characterizing the Impact of Active Queue Management on Speed Test Measurements
Siddhant Ray, Taveesh Sharma, Jonatas Marques +3
Present day speed test tools measure peak throughput, but often fail to capture the user-perceived responsiveness of a network connection under load. Recently, platforms such as ND…
SwiftQueue: Optimizing Low-Latency Applications with Swift Packet Queuing
Siddhant Ray, Xi Jiang, Jack Luo +2
Low Latency, Low Loss, and Scalable Throughput (L4S), as an emerging router-queue management technique, has seen steady deployment in the industry. An L4S-enabled router assigns ea…
Algorithmic Data Minimization for Machine Learning over Internet-of-Things Data Streams
Ted Shaowang, Shinan Liu, Jonatas Marques +2
Machine learning can analyze vast amounts of data generated by IoT devices to identify patterns, make predictions, and enable real-time decision-making. By processing sensor data,…
NetSSM: Multi-Flow and State-Aware Network Trace Generation using State Space Models
Andrew Chu, Xi Jiang, Shinan Liu +4
Access to raw network traffic data is essential for many computer networking tasks, from traffic modeling to performance evaluation. Unfortunately, this data is scarce due to high…
WiFinger: Fingerprinting Noisy IoT Event Traffic Using Packet-level Sequence Matching
Ronghua Li, Shinan Liu, Haibo Hu +2
IoT environments such as smart homes are susceptible to privacy inference attacks, where attackers can analyze patterns of encrypted network traffic to infer the state of devices a…