-
Aozorabunko dataset
Aozorabunko dataset used for pre-training of PnG BERT model. -
Wikipedia2 and Aozorabunko datasets
Wikipedia2 and Aozorabunko datasets used for pre-training of PnG BERT model. -
FastSpeech
The FastSpeech dataset is a text-to-speech dataset used to train the FastSpeech model. -
Style Tokens
Global Style Tokens (GSTs) are a recently-proposed method to learn latent disentangled representations of high-dimensional data. GSTs can be used within Tacotron, a... -
Global Style Tokens
Global Style Tokens (GSTs) are a recently-proposed method to learn latent disentangled representations of high-dimensional data. GSTs can be used within Tacotron, a... -
Text-Predicted Global Style Tokens
Global Style Tokens (GSTs) are a recently-proposed method to learn latent disentangled representations of high-dimensional data. GSTs can be used within Tacotron, a... -
LJ Speech Dataset
The LJ speech dataset is a dataset of speech samples recorded from a single speaker reading passages from 7 non-fiction books.