2 papers
quant-ph2025
Multi-Armed Bandits and Quantum Channel Oracles
Simon Buchholz, Jonas M. Kübler, Bernhard Schölkopf
Multi-armed bandits are one of the theoretical pillars of reinforcement learning. Recently, the investigation of quantum algorithms for multi-armed bandit problems was started, and…
cs.LG2024
Limits of Transformer Language Models on Learning to Compose Algorithms
Jonathan Thomm, Giacomo Camposampiero, Aleksandar Terzic +3
We analyze the capabilities of Transformer language models in learning compositional discrete tasks. To this end, we evaluate training LLaMA models and prompting GPT-4 and Gemini o…