2 papers
cs.LG2026
A Geometric Perspective on Next-Token Prediction in Large Language Models: Three Emerging Phases
Gianfranco Lombardo, Giuseppe Trimigno, Stefano Cagnoni
We investigate the geometry of predictive information across the layers of large language models (LLMs). We repurpose representation lenses-learned affine maps trained to predict t…
cs.NE2024
Evolutionary Computation and Explainable AI: A Roadmap to Understandable Intelligent Systems
Ryan Zhou, Jaume Bacardit, Alexander Brownlee +7
Artificial intelligence methods are being increasingly applied across various domains, but their often opaque nature has raised concerns about accountability and trust. In response…