9 citations · 9 across the 1 of their papers we have counts for
3 papers
cs.LG2025
Listwise Direct Preference Optimization with Multi-Dimensional Preference Mixing
Yuhui Sun, Xiyao Wang, Zixi Li +6
Recent alignment methods based on Direct Preference Optimization (DPO) reformulate preference learning as supervised optimization over pairwise comparisons, offering improved effic…
cs.CL2024
Role-Play Paradox in Large Language Models: Reasoning Performance Gains and Ethical Dilemmas
Jinman Zhao, Zifan Qian, Linbo Cao +5
Role-play in large language models (LLMs) enhances their ability to generate contextually relevant and high-quality responses by simulating diverse cognitive perspectives. However,…
cs.CL2024★ 9 cited
Gender Bias in Large Language Models across Multiple Languages
Jinman Zhao, Yitian Ding, Chen Jia +2
With the growing deployment of large language models (LLMs) across various applications, assessing the influence of gender biases embedded in LLMs becomes crucial. The topic of gen…