Showing cs.CLShow all
3 papers · 1 filter
cs.CL2026
Painless Activation Steering: An Automated, Lightweight Approach for Post-Training Large Language Models
Sasha Cui, Zhongren Chen
Language models (LMs) are typically post-trained for desired capabilities and behaviors via weight-based or prompt-based steering, but the former is time-consuming and expensive, a…
cs.CL2026
Benchmarking Political Persuasion Risks Across Frontier Large Language Models
Zhongren Chen, Joshua Kalla, Quan Le
Concerns persist regarding the capacity of Large Language Models (LLMs) to sway political views. Although prior research has claimed that LLMs are not more persuasive than standard…
cs.CL2025
A Framework to Assess the Persuasion Risks Large Language Model Chatbots Pose to Democratic Societies
Zhongren Chen, Joshua Kalla, Quan Le +3
In recent years, significant concern has emerged regarding the potential threat that Large Language Models (LLMs) pose to democratic societies through their persuasive capabilities…