2 papers
cs.CR2025
NeuroBreak: Unveil Internal Jailbreak Mechanisms in Large Language Models
Chuhan Zhang, Ye Zhang, Bowen Shi +5
Jailbreak attacks bypass the safety alignment of large language models (LLMs) to elicit harmful outputs, yet the vast parameter space makes diagnosing the underlying failure mechan…
cs.HC2025
NoteFlow: Leveraging Charts as Sight Glasses for Consistent and Continuous Data Flow Tracing
Yuan Tian, Dazhen Deng, Sen Yang +5
Computational notebooks offer a flexible environment for exploratory data analysis (EDA), but this flexibility often leads to disorganized and iterative execution of notebook cells…