43 papers
A Resource-centric Analysis and Optimization of NoSQL Workloads using Distressed Resource Volume Metric
Gunika Verma, Aashutosh A, Pooja Srinivas +9
Large-scale managed cloud databases leverage sophisticated load Packing and Migration (PAM) algorithms, which provide the efficiencies necessary for running these services at scale…
A 6G Integrated Sensing and Communication Framework for Railway Intrusion Detection and Collision Prediction
Ajeet Kumar Yadav, Sankaran Balasubramaniam, Aritra Chatterjee +3
Integrated Sensing and Communication (ISAC) combines sensing and communication to efficiently utilize wireless resources and is emerging as a key paradigm for next-generation wirel…
Taurus: Accelerating Out-of-Core Graph Neural Network Inference on Billion-Scale Graphs
Pranjal Naman, Yogesh Simmhan
Graph Neural Network (GNN) inference on billion-scale graphs is challenging due to the large memory footprint of features and embeddings and high disk I/O costs in out-of-core sett…
Collaborative Processing for Multi-Tenant Inference on Memory-Constrained Edge TPUs
Nathan Ng, Walid A. Hanafy, Prashanthi Kadambi +5
IoT applications increasingly rely on on-device AI accelerators to ensure high performance, especially in low-connectivity and safety-critical scenarios. However, the limited on-ch…
ATLAS: Efficient Out-of-Core Inference for Billion-Scale Graph Neural Networks
Pranjal Naman, Yogesh Simmhan
Graph Neural Network (GNN) inference on billion-scale graphs is critical for domains like fintech and recommendation systems. Full-graph inference on these large graphs can be chal…
Per-Phase Fidelity Attribution for Quantum Compilers using HBR Decomposition
Chandrachud Pati, Yogesh Simmhan
Quantum compilers sit between an algorithm's theoretical promise and what executes on physical hardware. Existing benchmarks report aggregate post-transpilation metrics but cannot…