1 citations · 1 across the 3 of their papers we have counts for
Showing cs.CLShow all
2 papers · 1 filter
cs.CL2024
A Hopfieldian View-based Interpretation for Chain-of-Thought Reasoning
Lijie Hu, Liang Liu, Shu Yang +6
Chain-of-Thought (CoT) holds a significant place in augmenting the reasoning performance for large language models (LLMs). While some studies focus on improving CoT accuracy throug…
cs.CL2024
Dialectical Alignment: Resolving the Tension of 3H and Security Threats of LLMs
Shu Yang, Jiayuan Su, Han Jiang +5
With the rise of large language models (LLMs), ensuring they embody the principles of being helpful, honest, and harmless (3H), known as Human Alignment, becomes crucial. While exi…