4 papers
Conformal Sparsification for Bandwidth-Efficient Edge-Cloud Speculative Decoding
Payel Bhattacharjee, Fengwei Tian, Meiyu Zhong +3
Edge-cloud speculative decoding (SD) accelerates inference by having a cloud-based large language model (LLM) that verifies draft tokens generated by a resource-constrained small l…
PROPS: Progressively Private Self-alignment of Large Language Models
Noel Teku, Fengwei Tian, Payel Bhattacharjee +3
Alignment is a key step in developing Large Language Models (LLMs) using human feedback to ensure adherence to human values and societal norms. Dependence on human feedback raises…
Inference Privacy: Properties and Mechanisms
Fengwei Tian, Ravi Tandon
Ensuring privacy during inference stage is crucial to prevent malicious third parties from reconstructing users' private inputs from outputs of public models. Despite a large body…
CURATE: Scaling-up Differentially Private Causal Graph Discovery
Payel Bhattacharjee, Ravi Tandon
Causal Graph Discovery (CGD) is the process of estimating the underlying probabilistic graphical model that represents joint distribution of features of a dataset. CGD-algorithms a…