1 paper
Bradley C. A. Brown, Jordan Juravsky, Anthony L. Caterini +1
Given a pair of models with similar training set performance, it is natural to assume that the model that possesses simpler internal representations would exhibit better generaliza…