AI product Open source
Moonshine Voice is an open-source AI toolkit for developers building real-time voice agents and applications. It provides on-device speech-to-text, intent recognition, and text-to-speech, including streaming transcription that processes audio while the user is still speaking. The project offers speech-to-text models ranging from higher-accuracy models to approximately 1 MB models and provides one library for Python, JavaScript/WASM, iOS, Android, macOS, Linux, Windows, and Raspberry Pi. It is used by Gemma Translator for speech recognition and speech output, and can run without an account or API keys. The project is distributed under the MIT License, including its models by default, with legacy non-streaming models for non-English languages covered by a non-commercial Moonshine Community License.
1 use taken from transcripts — each links to the moment in the video.
Handles speech recognition and speech output for Gemma Translator.
1 in the library.