1 citations · 1 across the 3 of their papers we have counts for
3 papers
cs.CL2024
Dialectical Alignment: Resolving the Tension of 3H and Security Threats of LLMs
Shu Yang, Jiayuan Su, Han Jiang +5
With the rise of large language models (LLMs), ensuring they embody the principles of being helpful, honest, and harmless (3H), known as Human Alignment, becomes crucial. While exi…
cs.LG2024
Causal State Distillation for Explainable Reinforcement Learning
Wenhao Lu, Xufeng Zhao, Thilo Fryen +4
Reinforcement learning (RL) is a powerful technique for training intelligent agents, but understanding why these agents make specific decisions can be quite challenging. This lack…
cs.RO2023★ 1 cited
Accelerating Reinforcement Learning of Robotic Manipulations via Feedback from Large Language Models
Kun Chu, Xufeng Zhao, Cornelius Weber +2
Reinforcement Learning (RL) plays an important role in the robotic manipulation domain since it allows self-learning from trial-and-error interactions with the environment. Still,…