Parler-TTS
the six Ws · specification
Hugging Face, building on research by Dan Lyth and Simon King, released Parler-TTS.
Parler-TTS is a text-to-speech model whose voice style, pitch, and recording quality can be steered with a natural language description prompt.
Hosted on GitHub at huggingface/parler-tts and on the Hugging Face Hub.
Released in 2024, with Parler-TTS Mini and Large versions following the original paper.
It gives developers fine-grained, prompt-based control over synthetic voice characteristics without recording new training data per voice.
Built on PyTorch and integrates with the Hugging Face transformers and datasets libraries.
Released under the Apache 2.0 license by Hugging Face; fully open weights.
alternatives
works with