1 paper
Yi-Hang Zhu, Rajeev Raman, Shiqi Su +4
Models trained on long-tailed data using standard softmax tend to exhibit higher training error and a larger generalisation gap for classes with fewer training samples. We characte…