3 papers
cs.AR2025
SPAD: Specialized Prefill and Decode Hardware for Disaggregated LLM Inference
Hengrui Zhang, Pratyush Patel, August Ning +1
Large Language Models (LLMs) have gained popularity in recent years, driving up the demand for inference. LLM inference is composed of two phases with distinct characteristics: a c…
cs.SE2025
Synthesizing Access Control Policies using Large Language Models
Adarsh Vatsa, Pratyush Patel, William Eiers
Cloud compute systems allow administrators to write access control policies that govern access to private data. While policies are written in convenient languages, such as AWS Iden…
cs.AI2024
Input-Dependent Power Usage in GPUs
Theo Gregersen, Pratyush Patel, Esha Choukse
GPUs are known to be power-hungry, and due to the boom in artificial intelligence, they are currently the major contributors to the high power demands of upcoming datacenters. Most…