2 papers
cs.CL2025
Hold Onto That Thought: Assessing KV Cache Compression On Reasoning
Minghui Liu, Aadi Palnitkar, Tahseen Rabbani +9
Large language models (LLMs) have demonstrated remarkable performance on long-context tasks, but are often bottlenecked by memory constraints. Namely, the KV cache, which is used t…
cs.LG2024
Balancing Label Imbalance in Federated Environments Using Only Mixup and Artificially-Labeled Noise
Kyle Sang, Tahseen Rabbani, Furong Huang
Clients in a distributed or federated environment will often hold data skewed towards differing subsets of labels. This scenario, referred to as heterogeneous or non-iid federated…