Supertonic
On-device ONNX text-to-speech around 99M parameters covering 31 languages with no cloud calls.
About
Supertonic goes where big TTS models cannot: about 99M parameters running through ONNX Runtime entirely on-device, synthesizing 44.1kHz audio faster than real time on plain CPUs, phones, browsers via WebGPU, and even Raspberry Pi boards and e-readers. Version 3, released May 2026, expanded coverage from 5 to 31 languages and ships ten preset voice styles plus inline expression tags for things like laughter and breaths. Integration is the selling point: ready-made SDKs cover Python, Node.js, browser JavaScript, Java, C++, C#, Go, Rust, Swift, and Flutter, a pip package installs the engine in one line, and a local HTTP server exposes OpenAI-compatible endpoints. Sample code is MIT while the model weights carry an OpenRAIL-M license, with everything downloadable from Hugging Face. One caveat: developer Supertone ended official support in July 2026 and its hosted Voice Builder is being retired, so the repo is frozen in a working state rather than evolving. For private, offline speech on weak hardware it remains hard to beat.
Should you use Supertonic?
Pick it when
Pick Supertonic when speech must be generated fully offline on weak hardware, such as phones, browsers, Raspberry Pi boards, or e-readers, in any of 31 languages, and you want ready SDKs for your app's language plus a local OpenAI-compatible server.
Look elsewhere when
Skip it if you need a project that will keep improving or cloning from your own samples: Supertone ended official support in July 2026 and is retiring its hosted Voice Builder. The OpenRAIL-M weight license also carries use restrictions.
Alternatives to Supertonic
- Kokoro TTS
82M Apache-2.0 model on CPU or GPU with streaming, fewer languages than Supertonic's 31 but a plainer license and an open Python library.
- Piper TTS
Offline VITS-based voices in dozens of languages that run on a Raspberry Pi 4, under GPL, with documented training for custom voices.
- KittenTTS
Even smaller at 15M to 80M parameters for the tiniest devices, but English only and still flagged as a developer preview.
- Sherpa-ONNX TTS
Toolkit that runs several ONNX TTS models across mobile, desktop, and embedded targets, so you are not tied to one frozen model.
Reviews (0)
Leave a Review
No reviews yet. Be the first to review!
Details
- Category
- Text-to-Speech (TTS)
- Price
- Free
- Platform
- Local/Desktop
- Difficulty
- Easy (2/5)
- License
- MIT / OpenRAIL-M
- Added
- Aug 24, 2026