3 papers
cs.CL2026
Kathleen Remembers: Length-Invariant One-Shot Recall Without Attention
George Fountzoulas
Recurrent, attention-free sequence models share a structural weakness: a fading state cannot perform exact recall of something seen once, far in the past. We add to the Kathleen tr…
cs.CL2026
Kathleen Writes: Autoregressive Generation and Data Scaling Without Attention
George Fountzoulas
Papers 1-2 of the Kathleen series showed that a byte-level, attention-free architecture built from a wavetable encoder and multi-scale reverberant state can match strong baselines…
cs.CL2026
Kathleen: Oscillator-Based Byte-Level Text Classification Without Tokenization or Attention
George Fountzoulas
We present Kathleen, a text classification architecture that operates directly on raw UTF-8 bytes using frequency-domain processing -- requiring no tokenizer, no attention mechanis…