moonshine-ai/moonshine
原文摘要
Very low latency speech to text, intent recognition, and text to speech, for building voice agents and interfaces Moonshine Voice Voice Interfaces for Everyone Quickstart When should you choose Moonshine over Whisper? Using the Library Speech to Text Text to Speech Conversational Agents Models API Reference Support Roadmap Acknowledgements License Moonshine Voice is an open source AI toolkit for developers building real-time voice agents and applications. Everything runs on-device, so it's fast, private, and you don't need an account, credit card, or API keys. The framework and models are optimized for live streaming applications, offering low latency responses by doing a lot of the work while the user is still talking. All speech to text models are based on our cutting edge research and trained from scratch, so we can offer higher accuracy than Whisper Large V3 at the top end , down to tiny 1MB models for constrained deployments . It's easy to integrate across platforms, with the same library running on Python , iOS , Android , MacOS , Linux , Windows , Raspberry Pis , IoT devices , microcontrollers , DSPs , and wearables. Batteries are included. Its high-level APIs offer complete…