3 citations · 3 across the 1 of their papers we have counts for
1 paper
Moran Mizrahi, Guy Kaplan, Dan Malkin +3
Recent advances in large language models (LLMs) have led to the development of various evaluation benchmarks. These benchmarks typically rely on a single instruction template for e…