Multispeaker & Emotional TTS based on Tacotron 2 and Waveglow
-
Updated
Apr 9, 2021 - Jupyter Notebook
Multispeaker & Emotional TTS based on Tacotron 2 and Waveglow
Implementation of Global Style Token Tacotron in TensorFlow2
ParsiGoo is a Persian multispeaker dataset for text-to-speech purposes. It includes recordings from different speakers and is designed to be used for training and evaluating text-to-speech models.
This is PyTorch Implementation of A Non-Autoregressive Transformer with unsupervised learning durations based on Transformer & Conformer blocks, supporting for Vietnamese language.
Distributed feedback system in Max, on its way to becoming a M4L effect...
Leverages the VIVOS corpus to synthesis two-speaker scenarios and fine-tunes the Conv-TasNet architecture for Vietnamese speech separation.
To associate your repository with the multispeaker topic, visit your repo's landing page and select "manage topics."