2 papers
cs.AR2025
SPAD: Specialized Prefill and Decode Hardware for Disaggregated LLM Inference
Hengrui Zhang, Pratyush Patel, August Ning +1
Large Language Models (LLMs) have gained popularity in recent years, driving up the demand for inference. LLM inference is composed of two phases with distinct characteristics: a c…
cs.SE2025
Synthesizing Access Control Policies using Large Language Models
Adarsh Vatsa, Pratyush Patel, William Eiers
Cloud compute systems allow administrators to write access control policies that govern access to private data. While policies are written in convenient languages, such as AWS Iden…