4 citations · 4 across the 3 of their papers we have counts for
3 papers · 1 filter
Shared Global and Local Geometry of Language Model Embeddings
Andrew Lee, Melanie Weber, Fernanda Viégas +1
Researchers have recently suggested that models share common representations. In our work, we find numerous geometric similarities across the token embeddings of large language mod…
ICLR: In-Context Learning of Representations
Core Francisco Park, Andrew Lee, Ekdeep Singh Lubana +5
Recent work has demonstrated that semantics specified by pretraining data influence how representations of different concepts are organized in a large language model (LLM). However…
Some things are more CRINGE than others: Iterative Preference Optimization with the Pairwise Cringe Loss
Jing Xu, Andrew Lee, Sainbayar Sukhbaatar +1
Practitioners commonly align large language models using pairwise preferences, i.e., given labels of the type response A is preferred to response B for a given input. Perhaps less…