17 citations · 17 across the 5 of their papers we have counts for
4 papers · 1 filter
Memory Inception: Latent-Space KV Cache Manipulation for Steering LLMs
Andy Zeyi Liu, Michael Zhang, Ilana Greenberg +3
Steering large language models (LLMs) is usually done by either instruction prompting or activation steering. Prompting often gives strong control, but caches guidance tokens at ev…
Fast Exact Unlearning for In-Context Learning Data for LLMs
Andrei I. Muresanu, Anvith Thudi, Michael R. Zhang +1
Modern machine learning models are expensive to train, and there is a growing concern about the challenge of retroactively removing specific training data. Achieving exact unlearni…
Using Large Language Models for Hyperparameter Optimization
Michael R. Zhang, Nishkrit Desai, Juhan Bae +2
This paper explores the use of foundational large language models (LLMs) in hyperparameter optimization (HPO). Hyperparameters are critical in determining the effectiveness of mach…
Report Cards: Qualitative Evaluation of Language Models Using Natural Language Summaries
Blair Yang, Fuyang Cui, Keiran Paster +4
The rapid development and dynamic nature of large language models (LLMs) make it difficult for conventional quantitative benchmarks to accurately assess their capabilities. We prop…