2 papers
cs.CL2024
SEER: Facilitating Structured Reasoning and Explanation via Reinforcement Learning
Guoxin Chen, Kexin Tang, Chao Yang +3
Elucidating the reasoning process with structured explanations from question to answer is crucial, as it significantly enhances the interpretability, traceability, and trustworthin…
cs.CL2024
Imposter.AI: Adversarial Attacks with Hidden Intentions towards Aligned Large Language Models
Xiao Liu, Liangzhi Li, Tong Xiang +4
With the development of large language models (LLMs) like ChatGPT, both their vast applications and potential vulnerabilities have come to the forefront. While developers have inte…