Parallel WaveGAN is an unofficial PyTorch implementation of several state-of-the-art non-autoregressive neural vocoders, centered on Parallel WaveGAN but also including MelGAN, Multiband-MelGAN, HiFi-GAN, and StyleMelGAN. Its main goal is to provide a real-time neural vocoder that can turn mel spectrograms into high-quality speech audio efficiently. The repository is designed to work hand-in-hand with ESPnet-TTS and NVIDIA Tacotron2-style front ends, so you can build complete TTS or singing voice synthesis pipelines. It includes a large collection of “Kaldi-style” recipes for many datasets such as LJSpeech, LibriTTS, VCTK, JSUT, CMU Arctic, and multiple singing voice corpora in Japanese, Mandarin, Korean, and more. The project provides pre-trained models, Colab demos, and example configurations, allowing researchers to quickly evaluate vocoder quality or adapt models to new datasets.

Features

  • PyTorch implementations of Parallel WaveGAN, MelGAN, Multiband-MelGAN, HiFi-GAN, and StyleMelGAN
  • Real-time neural vocoder compatible with ESPnet-TTS and Tacotron2 front ends
  • Extensive set of Kaldi-style recipes for speech and singing datasets in multiple languages
  • Pretrained models and Colab demos for quick listening tests and prototyping
  • Flexible training pipeline with support for multi-GPU and distributed setups
  • Very low real-time factor for fast mel-to-waveform conversion suitable for deployment

Project Samples

Project Activity

See All Activity >

Categories

Text to Speech

License

MIT License

Follow Parallel WaveGAN

Parallel WaveGAN Web Site

Other Useful Business Software
Our Free Plans just got better! | Auth0 Icon
Our Free Plans just got better! | Auth0

With up to 25k MAUs and unlimited Okta connections, our Free Plan lets you focus on what you do best—building great apps.

You asked, we delivered! Auth0 is excited to expand our Free and Paid plans to include more options so you can focus on building, deploying, and scaling applications without having to worry about your security. Auth0 now, thank yourself later.
Try free now
Rate This Project
Login To Rate This Project

User Reviews

Be the first to post a review of Parallel WaveGAN!

Additional Project Details

Programming Language

Python

Related Categories

Python Text to Speech Software

Registered

2025-11-28