3 papers
cs.CL2025
Machine Learning for Detection and Analysis of Novel LLM Jailbreaks
John Hawkins, Aditya Pramar, Rodney Beard +1
Large Language Models (LLMs) suffer from a range of vulnerabilities that allow malicious users to solicit undesirable responses through manipulation of the input text. These so-cal…
cs.AI2025
Improving AGI Evaluation: A Data Science Perspective
John Hawkins
Evaluation of potential AGI systems and methods is difficult due to the breadth of the engineering goal. We have no methods for perfect evaluation of the end state, and instead mea…
cs.AI2025
Enigme: Generative Text Puzzles for Evaluating Reasoning in Language Models
John Hawkins
Transformer-decoder language models are a core innovation in text based generative artificial intelligence. These models are being deployed as general-purpose intelligence systems…