5 citations · 6 across the 3 of their papers we have counts for
3 papers
cs.LG2023★ 5 cited
Do Not Blindly Imitate the Teacher: Using Perturbed Loss for Knowledge Distillation
Rongzhi Zhang, Jiaming Shen, Tianqi Liu +4
Knowledge distillation is a popular technique to transfer knowledge from large teacher models to a small student model. Typically, the student learns to imitate the teacher by mini…
cs.CL2023★ 1 cited
"Why is this misleading?": Detecting News Headline Hallucinations with Explanations
Jiaming Shen, Jialu Liu, Dan Finnie +3
Automatic headline generation enables users to comprehend ongoing news events promptly and has recently become an important task in web mining and natural language processing. With…
cs.LG2022
Data-Efficient Information Extraction from Form-Like Documents
Beliz Gunel, Navneet Potti, Sandeep Tata +3
Automating information extraction from form-like documents at scale is a pressing need due to its potential impact on automating business workflows across many industries like fina…