most citedLeveraging Domain Knowledge for Efficient Reward Modelling in RLHF: A Case-Study in E-Commerce Opinion Summarization

1 citations · 1 across the 3 of their papers we have counts for

collaborators
Showing cs.CLShow all

4 papers · 1 filter