4 papers
Controlling Multimodal Conversational Agents with Coverage-Enhanced Latent Actions
Yongqi Li, Hao Lang, Tieyun Qian +1
Vision-language models are increasingly employed as multimodal conversational agents (MCAs) for diverse conversational tasks. Recently, reinforcement learning (RL) has been widely…
Understanding Generalization in Role-Playing Models via Information Theory
Yongqi Li, Hao Lang, Fei Huang +2
Role-playing models (RPMs) are widely used in real-world applications but underperform when deployed in the wild. This degradation can be attributed to distribution shifts, includi…
Selective Weak-to-Strong Generalization
Hao Lang, Fei Huang, Yongbin Li
Future superhuman models will surpass the ability of humans and humans will only be able to \textit{weakly} supervise superhuman models. To alleviate the issue of lacking high-qual…
Debate Helps Weak-to-Strong Generalization
Hao Lang, Fei Huang, Yongbin Li
Common methods for aligning already-capable models with desired behavior rely on the ability of humans to provide supervision. However, future superhuman models will surpass the ca…