3 papers
cs.AI2026
DenseSteer: Steering Small Language Models towards Dense Math Reasoning
Yang Ouyang, Shuhang Lin, Jung-Eun Kim
Large language models (LLMs) demonstrate strong chain-of-thought (CoT) reasoning abilities, while smaller models (<= 3B parameters) significantly underperform on multi-step reasoni…
cs.CR2025
Layer-Level Self-Exposure and Patch: Affirmative Token Mitigation for Jailbreak Attack Defense
Yang Ouyang, Hengrui Gu, Shuhang Lin +6
As large language models (LLMs) are increasingly deployed in diverse applications, including chatbot assistants and code generation, aligning their behavior with safety and ethical…
cs.CL2025
Min-K%++: Improved Baseline for Detecting Pre-Training Data from Large Language Models
Jingyang Zhang, Jingwei Sun, Eric Yeats +5
The problem of pre-training data detection for large language models (LLMs) has received growing attention due to its implications in critical issues like copyright violation and t…