2 papers
cs.AI2026
Harmful Content Is Not Enough: Continuation Framing Moderates In-Context Emergent Misalignment
Peiyang Liu, Xi Wang, Ziqiang Cui +2
In-context learning (ICL) can induce emergent misalignment (EM), where narrow misaligned examples alter answers to unrelated questions. Existing prompts, however, conflate harmful-…
cs.AI2026
Learning from Contrasts: Synthesizing Reasoning Paths from Diverse Search Trajectories
Peiyang Liu, Zhirui Chen, Xi Wang +4
Monte Carlo Tree Search (MCTS) has been widely used for automated reasoning data exploration, but current supervision extraction methods remain inefficient. Standard approaches ret…