32 citations · 86 across the 18 of their papers we have counts for
Showing 2023Show all
2 papers · 1 filter
cs.CL2023
ContextRef: Evaluating Referenceless Metrics For Image Description Generation
Elisa Kreiss, Eric Zelikman, Christopher Potts +1
Referenceless metrics (e.g., CLIPScore) use pretrained vision--language models to assess image descriptions directly without costly ground-truth reference texts. Such methods can f…
cs.CL2023
Rigorously Assessing Natural Language Explanations of Neurons
Jing Huang, Atticus Geiger, Karel D'Oosterlinck +2
Natural language is an appealing medium for explaining how large language models process and store information, but evaluating the faithfulness of such explanations is challenging.…