Showing cs.CLShow all
3 papers · 1 filter
cs.CL2026
Controllable and explainable personality sliders for LLMs at inference time
Florian Hoppe, David Khachaturov, Robert Mullins +1
Aligning Large Language Models (LLMs) with specific personas typically relies on expensive and monolithic Supervised Fine-Tuning (SFT) or RLHF. While effective, these methods requi…
cs.CL2025
Detecting Manipulated Contents Using Knowledge-Grounded Inference
Mark Huasong Meng, Ruizhe Wang, Meng Xu +2
The detection of manipulated content, a prevalent form of fake news, has been widely studied in recent years. While existing solutions have been proven effective in fact-checking a…
cs.CL2024
GlitchProber: Advancing Effective Detection and Mitigation of Glitch Tokens in Large Language Models
Zhibo Zhang, Wuxia Bai, Yuxi Li +6
Large language models (LLMs) have achieved unprecedented success in the field of natural language processing. However, the black-box nature of their internal mechanisms has brought…