Showing cs.AIShow all
3 papers · 1 filter
cs.AI2025
Controllable Mathematical Reasoning via Self-Optimizing Thought Vectors
Xuying LI
We present a novel approach for controllable mathematical reasoning that leverages self-optimizing thought vectors with entropy minimization. Our method introduces learnable though…
cs.AI2025
LatentGuard: Controllable Latent Steering for Robust Refusal of Attacks and Reliable Response Generation
Huizhen Shu, Xuying Li, Zhuo Li
Achieving robust safety alignment in large language models (LLMs) while preserving their utility remains a fundamental challenge. Existing approaches often struggle to balance comp…
cs.AI2024
Targeting the Core: A Simple and Effective Method to Attack RAG-based Agents via Direct LLM Manipulation
Xuying Li, Zhuo Li, Yuji Kosuga +2
AI agents, powered by large language models (LLMs), have transformed human-computer interactions by enabling seamless, natural, and context-aware communication. While these advance…