PaddleSpeech is an open-source toolkit on PaddlePaddle platform for a variety of critical tasks in speech and audio, with state-of-art and influential models. Via the easy-to-use, efficient, flexible and scalable implementation, our vision is to empower both industrial application and academic research, including training, inference & testing modules, and deployment process. Low barriers to install, CLI, Server, and Streaming Server is available to quick-start your journey. We provide high-speed and ultra-lightweight models, and also cutting-edge technology. We provide production ready streaming asr and streaming tts system. Our frontend contains Text Normalization and Grapheme-to-Phoneme (G2P, including Polyphone and Tone Sandhi). Moreover, we use self-defined linguistic rules to adapt Chinese context.

Features

  • Implementation of critical audio tasks
  • Integration of mainstream models and datasets
  • Cascaded models application
  • Streaming ASR and TTS System
  • Rule-based Chinese frontend
  • The toolkit implements modules that participate in the whole pipeline of the speech tasks

Project Samples

Project Activity

See All Activity >

License

Apache License V2.0

Follow PaddleSpeech

PaddleSpeech Web Site

Other Useful Business Software
Build Securely on AWS with Proven Frameworks Icon
Build Securely on AWS with Proven Frameworks

Lay a foundation for success with Tested Reference Architectures developed by Fortinet’s experts. Learn more in this white paper.

Moving to the cloud brings new challenges. How can you manage a larger attack surface while ensuring great network performance? Turn to Fortinet’s Tested Reference Architectures, blueprints for designing and securing cloud environments built by cybersecurity experts. Learn more and explore use cases in this white paper.
Download Now
Rate This Project
Login To Rate This Project

User Reviews

Be the first to post a review of PaddleSpeech!

Additional Project Details

Operating Systems

Mac, Windows

Programming Language

C++, Python

Related Categories

Python Voice Cloning Software, Python LLM Inference Tool, C++ Voice Cloning Software, C++ LLM Inference Tool

Registered

2023-03-23