Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
How Many Different Outputs Can a Transformer Generate?
Maxime Meyer, Mario Michelessa, Caroline Chaux +1
We study how we can leverage only a handful of characteristics of a transformer's architecture to closely predict the number of different sequences it can output, both qualitativel…
cs.LG2025
Memory Limitations of Prompt Tuning in Transformers
Maxime Meyer, Mario Michelessa, Caroline Chaux +1
Despite the empirical success of prompt tuning in adapting pretrained language models to new tasks, theoretical analyses of its capabilities remain limited. Existing theoretical wo…