1 paper
Wesley Brewer, Murali Meena Gopalakrishnan, Matthias Maiterth +12
With the end of Moore's law and Dennard scaling, efficient training increasingly requires rethinking data volume. Can we train better models with significantly less data via intell…