6 citations · 25 across the 12 of their papers we have counts for
Showing 2018Show all
2 papers · 1 filter
cs.CL2018
Counterfactual Learning from Human Proofreading Feedback for Semantic Parsing
Carolin Lawrence, Stefan Riezler
In semantic parsing for question-answering, it is often too expensive to collect gold parses or even gold answers as supervision signals. We propose to convert model outputs into a…
cs.CL2018
Improving a Neural Semantic Parser by Counterfactual Learning from Human Bandit Feedback
Carolin Lawrence, Stefan Riezler
Counterfactual learning from human bandit feedback describes a scenario where user feedback on the quality of outputs of a historic system is logged and used to improve a target sy…