2 citations · 2 across the 5 of their papers we have counts for
3 papers · 1 filter
A Generalized Optimization Engine (GOE) for Edge AI Inference Acceleration
Venkat R. Dasari, Jakob A. Adams, Vinod K. Mishra +1
Artificial intelligence (AI) models have demonstrated remarkable capabilities across various domains, yet their widespread deployment is impeded by significant computational costs,…
Transferable Latency Prediction for Fast LLM Screening on Heterogeneous Edge Devices
Xiaolong Tu, Vinod K. Mishra, Venkat R. Dasari +2
Accurate latency prediction is critical for deploying large language models (LLMs) on heterogeneous edge devices, where inference latency is affected by model architecture, prompt…
Optimization problems with low SWaP tactical Computing
Mee Seong Im, Venkat R. Dasari, Lubjana Beshaj +1
In a resource-constrained, contested environment, computing resources need to be aware of possible size, weight, and power (SWaP) restrictions. SWaP-aware computational efficiency…