Vibeleaderboard
Index / tool

SupertonicTTS Voice Style Extractor

github.com/kdrkdrkdr/supertonic.embed
Visit github.com
Category
AI Tools
Rank
No. 2311Tools index
Pricing
Open Source
Type
TOOL
GitHub
42 stars
Date

About

A tool that extracts a voice style embedding from any WAV recording for use with SupertonicTTS, producing a style JSON compatible with the model's shipped voice presets even though the original style encoder was never publicly released. It requires an NVIDIA GPU, needing about 10GB by default or 2.5GB with a reduced batch setting, and ships with a demo of five example voices synthesized in five languages.

Why it made the leaderboard

Unlocks custom voice cloning for SupertonicTTS users who were previously limited to its five shipped preset voices, filling a gap the original project left open.

Tags

ttsvoice-cloningsupertonicttsvoice-styleresearch

Tech Stack

Python

Comments (0)

No comments yet

Editorially curated, with community endorsements as a secondary signal. Corrections welcome.