2 papers
cs.LG2024
Prepacking: A Simple Method for Fast Prefilling and Increased Throughput in Large Language Models
Siyan Zhao, Daniel Israel, Guy Van den Broeck +1
During inference for transformer-based large language models (LLM), prefilling is the computation of the key-value (KV) cache for input tokens in the prompt prior to autoregressive…
cs.AI2023
High Dimensional Causal Inference with Variational Backdoor Adjustment
Daniel Israel, Aditya Grover, Guy Van den Broeck
Backdoor adjustment is a technique in causal inference for estimating interventional quantities from purely observational data. For example, in medical settings, backdoor adjustmen…