1 paper
Kaixin Ma, Filip Ilievski, Jonathan Francis +3
Commonsense reasoning benchmarks have been largely solved by fine-tuning language models. The downside is that fine-tuning may cause models to overfit to task-specific data and the…