Generate expressive, multilingual speech from a short voice reference with an open local model.
Audio & Voice
Turn English speech into punctuated text with word timestamps using an NVIDIA open ASR model.
Best for: Builders and researchers who want an open transcription model rather than a subscription-only editor.
Generate expressive, multilingual speech from a short voice reference with an open local model.
A browser audio toolkit for text-to-speech, speech-to-text, vocal removal, voice enhancement, and quick editing.
Run multilingual speech recognition or translation with an open NVIDIA model.