Top AI Repos — open-source AI, indexed and scored
Top AI Repos tracks AI repositories on GitHub and answers two different questions about each one: is it moving right now, and would you bet a product on it.
Top AI Repos tracks AI repositories on GitHub and answers two different questions about each one: is it moving right now, and would you bet a product on it.
A PyTorch implementation of Style Tokens: Unsupervised Style Modeling, Control and Transfer in End-to-End Speech Synthesis
| Date | Stars |
|---|---|
| 2026-07-24 | 374 |
| 2026-07-25 | 374 |
| 2026-07-28 | 374 |
| 2026-07-30 | 374 |
| 2026-08-06 | 374 |
Today
— stars today
This week
— stars this week
This month
— stars this month
Momentum
0.0
growth rate 0.00%/day
# GST-Tacotron-Pytorch
A PyTorch implementation of [Style Tokens: Unsupervised Style Modeling, Control and Transfer in End-to-End Speech Synthesis](https://arxiv.org/abs/1803.09017)

## Update
Add support for blizzard dataset.
## Requirements
``` shell
pip3 install -r requirements.txt
```
## File structure
- `Hyperparameters.py` --- hyperparameters
- `Network.py` --- encoder and decoder
- `Modules.py` --- some modules for tacotron
- `Loss.py` --- loss function
- `Data.py` --- dataset loader
- `utils.py` --- some util functions for data I/O
- `Synthesis.py` --- speech generation
## How to train
- Download a multispeaker dataset
- Preprocess your data and implement your `get_XX_data` function in `Data.py`
- Set hyperparameters in `Hyperparameters.py`
- Make a directory named `log` as follow:
```
--- log
| |
| --- log[log_number]
|
--- code
|
--- Tacotron
|
--- train.py
|
--- Network.py
|
......
```
- Run train.py
``` shell
python3 train.py [log_number] [dataset_size] [start_epoch]
[log_number]: the log directory number
[dataset_size]: int or all
[start_epoch]: which epoch start to train (0 if start from scratch )
for example:
python3 train.py 0 all 0
```
## How to generate wav
Run`generate.py`. Replace the `text` in `generate.py` with any chinese sentences as you like before running
> The pretained model provided is trained on Chinese dataset, so it only supports chinese now.
## Star History
[](https://star-history.com/#KinglittleQ/GST-Tacotron&Date)
Excerpt of 1,705 characters
Read on GitHubWould you bet a product on this? Bounded 0–100 and slow moving.
matched fp:f156eada81e27e31, topic:tts, desc:speech synthesis, readme:speech synthesis
matched fp:f156eada81e27e31, topic:pytorch