2 papers
cs.CL2026
"I've Seen How This Goes": Characterizing Diversity via Progressive Conditional Surprise
Matthew Khoriaty, David Williams-King, Shi Feng
Measuring the diversity of creative outputs is central to evaluating post-training mode collapse, comparing decoding strategies, and quantifying creative behavior in both AI and hu…
cs.LG2025
Don't Forget It! Conditional Sparse Autoencoder Clamping Works for Unlearning
Matthew Khoriaty, Andrii Shportko, Gustavo Mercier +1
Recent developments in Large Language Model (LLM) capabilities have brought great potential but also posed new risks. For example, LLMs with knowledge of bioweapons, advanced chemi…