2 citations · 3 across the 2 of their papers we have counts for
3 papers · 1 filter
Learning to Attribute with Attention
Benjamin Cohen-Wang, Yung-Sung Chuang, Aleksander Madry
Given a sequence of tokens generated by a language model, we may want to identify the preceding tokens that influence the model to generate this sequence. Performing such token att…
ContextCite: Attributing Model Generation to Context
Benjamin Cohen-Wang, Harshay Shah, Kristian Georgiev +1
How do language models use information provided as context when generating a response? Can we infer whether a particular generated statement is actually grounded in the context, a…
Ask Your Distribution Shift if Pre-Training is Right for You
Benjamin Cohen-Wang, Joshua Vendrow, Aleksander Madry
Pre-training is a widely used approach to develop models that are robust to distribution shifts. However, in practice, its effectiveness varies: fine-tuning a pre-trained model imp…