107 citations · 165 across the 22 of their papers we have counts for
Showing 2026Show all
2 papers · 1 filter
cs.AI2026
Contextual Value Alignment via Multilayer Combinatorial Fusion
Yuanhong Wu, Djallel Bouneffouf, D. Frank Hsu
Aligning large language models (LLMs) with human values remains a major challenge, especially for trustworthy AI. While existing approaches such as RLHF, CAI, and their variants ha…
cs.MA2026
Enhancing Value Alignment of LLMs with Multi-agent system and Combinatorial Fusion
Yuanhong Wu, Djallel Bouneffouf, D. Frank Hsu
Aligning large language models (LLMs) with human values is a central challenge for ensuring trustworthy and safe deployment. While existing methods such as Reinforcement Learning f…