most citedLeveraging Domain Knowledge for Efficient Reward Modelling in RLHF: A Case-Study in E-Commerce Opinion Summarization

1 citations · 1 across the 4 of their papers we have counts for

collaborators

4 papers