Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
Voxtral TTS
Mistral-AI, :, Alexander H. Liu +186
We introduce Voxtral TTS, an expressive multilingual text-to-speech model that generates natural speech from as little as 3 seconds of reference audio. Voxtral TTS adopts a hybrid…
cs.AI2025
MindForge: Empowering Embodied Agents with Theory of Mind for Lifelong Cultural Learning
Mircea LicÄ, Ojas Shirekar, Baptiste Colle +1
Embodied agents powered by large language models (LLMs), such as Voyager, promise open-ended competence in worlds such as Minecraft. However, when powered by open-weight LLMs they…