It turns your text into high-quality speech and lets you customise voice profiles, speed and pitch.
Audio & Voice

A high-performance open-source text-to-speech model providing granular control over vocal characteristics and custom cloning.
Best for: ML engineers and digital content creators seeking locally-hosted, high-fidelity synthetic voice generation.
It turns your text into high-quality speech and lets you customise voice profiles, speed and pitch.
It helps you kokoro-82M TTS is an ultra-fast, open-source text-to-speech model that transforms your texts into natural speech (82M parameters only).
It helps you clone any voice in seconds and control the emotion of the speech maked.