- Separates slower linguistic planning from faster acoustic detail generation — meaning and waveform detail do not have to run at the same rate.
- Firefly-GAN decoder reconstructs the waveform from quantized features.
Connections
- part_of Text-to-Speech
- uses Firefly-GAN