2 papers
cs.CL2026
Multi-Turn Reflective Masking Elicits Reasoning in Mask Diffusion Models
Yanming Zhang, Yihan Bian, Jingyuan Qi +3
While reasoning on autoregressive (AR) models is often performed by chain-of-thought reasoning and reflection, their refinement of previous outputs still relies on fully sequential…
cs.LG2026
StagePilot: Stage-Level Planning for Long-Horizon Dialogue Simulation in Cybergrooming
Heajun An, Qi Zhang, Minqian Liu +5
Cybergrooming is an evolving threat to youth, requiring proactive educational interventions. We address this by modeling dialogue progression as a structured planning problem over…