supertonic

Lightning-Fast, On-Device, Multilingual TTS — running natively via ONNX.

88 articles 4.9k View on GitHub ↗
88 articles
How the Supertonic Speed Parameter (0.7–2.0) Affects Latency vs. Audio Quality

Discover how the Supertonic speed parameter (0.7-2.0) impacts audio latency and quality. Learn how to balance natural speech with computational speed for optimal results.

performance
Jun 15, 2026
How to Migrate from Supertonic v2 to v3 While Maintaining ONNX Interface Compatibility

Migrate Supertonic v2 to v3 seamlessly. Maintain ONNX interface compatibility by swapping model and config files. Learn how to upgrade without inference code changes.

migration-guide
Jun 15, 2026
Supertonic vs VoxCPM2 WER/CER: How the 99M Parameter Model Matches 2B Parameter Accuracy

Discover how Supertonic's 99M parameter TTS model rivals VoxCPM2's 2B parameter accuracy in WER/CER. Get comparable or better results with fewer resources.

performance
Jun 15, 2026
How Supertonic Preprocesses Text: Unicode Normalization and Emoji Removal

Discover how Supertonic preprocesses text with its UnicodeProcessor. Learn about Unicode normalization, emoji removal, and a 10-step pipeline for clean, tokenized input.

how-to-guide
Jun 15, 2026
How to Optimize Supertonic for Low‑Latency Real‑Time Synthesis on CPU

Optimize Supertonic for low-latency CPU synthesis. Configure ONNX Runtime, reduce steps to 6-8, boost speed to 1.2-1.5, and use singleton sessions for minimal overhead.

performance
Jun 15, 2026
How Supertonic’s ONNX Inference Pipeline Works: Duration Predictor, Text Encoder, Vector Estimator, and Vocoder Explained

Explore Supertonic's ONNX inference pipeline. Understand how duration predictor, text encoder, vector estimator, and vocoder work together for advanced text to speech generation.

internals
Jun 15, 2026
Required ONNX Model Files and How load_onnx_all() Loads Them in Supertonic

Discover the four essential ONNX model files required by Supertonic and learn how load_onnx_all() efficiently loads them using onnxruntime for your AI audio projects.

api-reference
Jun 14, 2026
How to Set Up the Supertonic Serve HTTP Server with Custom Endpoints

Learn to set up the supertonic serve HTTP server with custom endpoints. Extend built-in TTS functionality with your own business logic using FastAPI and APIRouter.

how-to-guide
Jun 14, 2026
Supertonic Text-to-Speech Output Audio Format: 44.1 kHz 16-bit WAV Configuration

Discover Supertonic's default output audio format: 44.1 kHz 16-bit WAV mono files. Configure your audio settings for high-fidelity text-to-speech. Optimize your sound projects with Supertonic.

api-reference
Jun 14, 2026
Maximum Input Text Length and Text Chunking in Supertonic: Implementation Guide

Discover Supertonic's maximum input text length limits and how it intelligently chunks long text into 300-character segments by default using sentence boundaries and a consistent algorithm.

how-to-guide
Jun 14, 2026
How Supertonic Handles Financial Expressions, Phone Numbers, and Technical Units

Discover how Supertonic's TTS pipeline automatically handles financial expressions, phone numbers, and technical units with its built-in text normalizer. Get natural spoken output without pre-processing.

deep-dive
Jun 14, 2026
How to Load and Switch Between Voice Style Presets (M1‑M5, F1‑F5) in Supertonic

Learn to load and switch Supertonic voice style presets M1-M5 and F1-F5 using loadVoiceStyle or load_voice_style. Easily manage voice styles at runtime by updating file paths or batch indexes.

how-to-guide
Jun 14, 2026
…

Have a question about this repo?

These articles cover the highlights, but your codebase questions are specific. Give your agent direct access to the source. Share this with your agent to get started:

Share the following with your agent to get started:
curl -s "https://instagit.com/install.md"

Works with
Claude Codex Cursor VS Code OpenClaw Any MCP Client

Maintain an open-source project? Get it listed too →