Kana & Mari's SoundRepos (English)
This repository provides a PyTorch implementation of a convolutional text-to-speech system based on the paper "Efficiently Trainable Text-to-Speech System Based on Deep Convolutional Networks with Guided Attention." It supports dataset downloading and audio preprocessing, training separate Text2Mel and SSRN spectrogram models, guided attention, checkpointing, TensorBoard logging, and speech synthesis for English LJ Speech and Mongolian Bible datasets.
45 afleveringen
Reacties
0Wees de eerste die een reactie plaatst
Meld je nu aan en word lid van de Kana & Mari's SoundRepos (English) community!