Showing cs.CLShow all
2 papers · 1 filter
cs.CL2025
Sparks of Science: Hypothesis Generation Using Structured Paper Data
Charles O'Neill, Tirthankar Ghosal, Roberta RÄileanu +4
Generating novel and creative scientific hypotheses is a cornerstone in achieving Artificial General Intelligence. Large language and reasoning models have the potential to aid in…
cs.CL2024
Sparse Autoencoders Enable Scalable and Reliable Circuit Identification in Language Models
Charles O'Neill, Thang Bui
This paper introduces an efficient and robust method for discovering interpretable circuits in large language models using discrete sparse autoencoders. Our approach addresses key…