Supertonic 3 Lightning Fast, On Device, Accurate TTS Supertonic is a lightweight text to speech system for local inference. It runs with ONNX Runtime entirely on your device, with no cloud call required for synthesis. Supertonic 3 expands the open weight release from 5 to 31 languages , improves reading stability, and reduces repeat/skip failures. Quick Start Install the Python SDK and generate speech immediately. On first run, the SDK downloads the model assets from Hugging Face. What's New in Supertonic 3 31 languages : expanded from the 5 language Supertonic 2 release. More stable reading : fewer repeat and skip failures, especially on short and long utterances. Higher speaker similarity : improved similarity across the shared language set compared with Supertonic 2. Expression tags : supports simple tags such as , , and . Custom Voices and Audio Samples The open weight package includes fixed preset voice styles for immediate local inference. If you want to hear how Supertonic 3 performs with zero shot custom voice styles, visit the Audio Sample Demo to compare reference audio and generated speech across several use cases. To create your own Supertonic 3 voice style JSON from re…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy