
SupertonicTTS Voice Style Extractor
github.com/kdrkdrkdr/supertonic.embed- Category
- AI Tools
- Rank
- No. 2311Tools index
- Pricing
- Open Source
- Type
- TOOL
- GitHub
- 42 stars
- Date
About
A tool that extracts a voice style embedding from any WAV recording for use with SupertonicTTS, producing a style JSON compatible with the model's shipped voice presets even though the original style encoder was never publicly released. It requires an NVIDIA GPU, needing about 10GB by default or 2.5GB with a reduced batch setting, and ships with a demo of five example voices synthesized in five languages.
Why it made the leaderboard
Unlocks custom voice cloning for SupertonicTTS users who were previously limited to its five shipped preset voices, filling a gap the original project left open.
Tags
ttsvoice-cloningsupertonicttsvoice-styleresearch
Tech Stack
Python
Comments (0)
No comments yet
Editorially curated, with community endorsements as a secondary signal. Corrections welcome.