Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
Which Pairs to Compare for LLM Post-Training?
Jiangze Han, Vineet Goyal, Will Ma
Preference-based post-training has become a central paradigm for aligning language models. A common data-collection strategy is to generate a small set of completions for each prom…
cs.AI2026
PREFER: Personalized Review Summarization with Online Preference Learning
Millend Roy, Agostino Capponi, Vineet Goyal
Product reviews significantly influence purchasing decisions on e-commerce platforms. However, the sheer volume of reviews can overwhelm users, obscuring the information most relev…