Tacotron is a heavily documented TensorFlow implementation of the end-to-end text-to-speech architecture introduced in the original Tacotron paper. It converts text into speech by learning acoustic representations and attention-based alignments from paired text and audio. The repository includes preprocessing, model modules, training, evaluation, and synthesis scripts. Example training setups use LJ Speech, Nick Offerman audiobook recordings, and the World English Bible dataset. Users can monitor loss and attention plots during training to evaluate alignment quality. Pretrained checkpoints and generated samples are provided as references for reproducing or studying the model's behavior.

Features

  • End-to-end text-to-speech synthesis
  • TensorFlow Tacotron implementation
  • Audio and text preprocessing
  • Training and evaluation scripts
  • Attention alignment visualization
  • Pretrained checkpoints and speech samples

Project Samples

Project Activity

See All Activity >

Categories

AI Models

License

Apache License V2.0

Follow tacotron

tacotron Web Site

Other Useful Business Software
MongoDB Atlas runs apps anywhere Icon
MongoDB Atlas runs apps anywhere

Deploy in 115+ regions with the modern database for every enterprise.

MongoDB Atlas gives you the freedom to build and run modern applications anywhere—across AWS, Azure, and Google Cloud. With global availability in over 115 regions, Atlas lets you deploy close to your users, meet compliance needs, and scale with confidence across any geography.
Start Free
Rate This Project
Login To Rate This Project

User Reviews

Be the first to post a review of tacotron!

Additional Project Details

Programming Language

Python

Related Categories

Python AI Models

Registered

5 days ago