Showing cs.AIShow all
3 papers · 1 filter
cs.AI2026
Shattering the Autoregressive Curse: Dynamic Epistemic Entropy Orchestrated Erasable Reinforcement Learning for LLMs
Ziliang Wang, Kang An, Faqiang Qian +5
Although reinforcement learning (RL) has expanded the cognitive boundaries of large language models (LLMs), it often remains vulnerable to the autoregressive curse in long-horizon…
cs.AI2026
SELF-EMO: Emotional Self-Evolution from Recognition to Consistent Expression
Shaowei Zhang, Faqiang Qian, Yan Chen +5
Emotion Recognition in Conversation (ERC) has become a fundamental capability for large language models (LLMs) in human-centric interaction. Beyond accurate recognition, coherent e…
cs.AI2025
AAPA: Adversarially Anchored Preference Alignment for Post-Training of Large Language Models
Faqiang Qian, Kang An, Weikun Zhang +6
Post-training alignment of large language models often combines supervised fine-tuning (SFT) on expert demonstrations with reinforcement learning (RL) from preference or verifiable…