Showing cs.LGShow all
2 papers · 1 filter
cs.LG2025
Depth-Wise Activation Steering for Honest Language Models
Gracjan Góral, Marysia Winkels, Steven Basart
Large language models sometimes assert falsehoods despite internally representing the correct answer, failures of honesty rather than accuracy, which undermines auditability and sa…
cs.LG2018
3D G-CNNs for Pulmonary Nodule Detection
Marysia Winkels, Taco S. Cohen
Convolutional Neural Networks (CNNs) require a large amount of annotated data to learn from, which is often difficult to obtain in the medical domain. In this paper we show that th…