1 paper
Subhojyoti Mukherjee, Anusha Lalitha, Sailik Sengupta +2
Multi-objective alignment from human feedback (MOAHF) in large language models (LLMs) is a challenging problem as human preferences are complex, multifaceted, and often conflicting…