4 citations · 4 across the 2 of their papers we have counts for
2 papers
cs.LG2023
Pushing the Accuracy-Group Robustness Frontier with Introspective Self-play
Jeremiah Zhe Liu, Krishnamurthy Dj Dvijotham, Jihyeon Lee +4
Standard empirical risk minimization (ERM) training can produce deep neural network (DNN) models that are accurate on average but under-perform in under-represented population subg…
cs.CL2023★ 4 cited
Understanding Finetuning for Factual Knowledge Extraction from Language Models
Mehran Kazemi, Sid Mittal, Deepak Ramachandran
Language models (LMs) pretrained on large corpora of text from the web have been observed to contain large amounts of various types of knowledge about the world. This observation h…