3 papers
cs.AI2025
From Model Training to Model Raising
Roland Aydin, Christian Cyron, Steve Bachelor +2
Current AI training methods align models with human values only after their core capabilities have been established, resulting in models that are easily misaligned and lack deep-ro…
cs.LG2025
One-Step is Enough: Sparse Autoencoders for Text-to-Image Diffusion Models
Viacheslav Surkov, Chris Wendler, Antonio Mari +5
For large language models (LLMs), sparse autoencoders (SAEs) have been shown to decompose intermediate representations that often are not interpretable directly into sparse sums of…
cs.CL2025
Fleet of Agents: Coordinated Problem Solving with Large Language Models
Lars Klein, Nearchos Potamitis, Roland Aydin +3
While numerous frameworks have been developed to enhance the reasoning abilities of large language models (LLMs), there is a scarcity of methods that effectively balance the trade-…