collaborators

5 papers

quant-ph2026

The Quantum Sieve Tracer: A Hybrid Framework for Layer-Wise Activation Tracing in Large Language Models

Jonathan Pan

Mechanistic interpretability aims to reverse-engineer the internal computations of Large Language Models (LLMs), yet separating sparse semantic signals from high-dimensional polyse…

cs.CL2026

Conversational Context Classification: A Representation Engineering Approach

Jonathan Pan

The increasing prevalence of Large Language Models (LLMs) demands effective safeguards for their operation, particularly concerning their tendency to generate out-of-context respon…

cs.CR2026

Automated Post-Incident Policy Gap Analysis via Threat-Informed Evidence Mapping using Large Language Models

Huan Lin Oh, Jay Yong Jun Jie, Mandy Lee Ling Siu +1

Cybersecurity post-incident reviews are essential for identifying control failures and improving organisational resilience, yet they remain labour-intensive, time-consuming, and he…

cs.LG2025

Probing Latent Subspaces in LLM for AI Security: Identifying and Manipulating Adversarial States

Xin Wei Chia, Swee Liang Wong, Jonathan Pan

Large Language Models (LLMs) have demonstrated remarkable capabilities across various tasks, yet they remain vulnerable to adversarial manipulations such as jailbreaking via prompt…

cs.CR2025

Automating Security Audit Using Large Language Model based Agent: An Exploration Experiment

Jia Hui Chin, Pu Zhang, Yu Xin Cheong +1

In the current rapidly changing digital environment, businesses are under constant stress to ensure that their systems are secured. Security audits help to maintain a strong securi…