3 papers
cs.CL2026
Moonshine v2: Ergodic Streaming Encoder ASR for Latency-Critical Speech Applications
Manjunath Kudlur, Evan King, James Wang +1
Latency-critical speech applications (e.g., live transcription, voice commands, and real-time translation) demand low time-to-first-token (TTFT) and high transcription accuracy, pa…
cs.CL2025
Flavors of Moonshine: Tiny Specialized ASR Models for Edge Devices
Evan King, Adam Sabra, Manjunath Kudlur +2
We present the Flavors of Moonshine, a suite of tiny automatic speech recognition (ASR) models specialized for a range of underrepresented languages. Prevailing wisdom suggests tha…
cs.SD2024
Moonshine: Speech Recognition for Live Transcription and Voice Commands
Nat Jeffries, Evan King, Manjunath Kudlur +3
This paper introduces Moonshine, a family of speech recognition models optimized for live transcription and voice command processing. Moonshine is based on an encoder-decoder trans…