PaddleSpeech is an open-source toolkit on PaddlePaddle platform for a variety of critical tasks in speech and audio, with state-of-art and influential models. Via the easy-to-use, efficient, flexible and scalable implementation, our vision is to empower both industrial application and academic research, including training, inference & testing modules, and deployment process. Low barriers to install, CLI, Server, and Streaming Server is available to quick-start your journey. We provide high-speed and ultra-lightweight models, and also cutting-edge technology. We provide production ready streaming asr and streaming tts system. Our frontend contains Text Normalization and Grapheme-to-Phoneme (G2P, including Polyphone and Tone Sandhi). Moreover, we use self-defined linguistic rules to adapt Chinese context.

Features

  • Implementation of critical audio tasks
  • Integration of mainstream models and datasets
  • Cascaded models application
  • Streaming ASR and TTS System
  • Rule-based Chinese frontend
  • The toolkit implements modules that participate in the whole pipeline of the speech tasks

Project Samples

Project Activity

See All Activity >

License

Apache License V2.0

Follow PaddleSpeech

PaddleSpeech Web Site

Other Useful Business Software
Cut Data Warehouse Costs by 54% Icon
Cut Data Warehouse Costs by 54%

Easily migrate from Snowflake, Redshift, or Databricks with free tools.

BigQuery delivers 54% lower TCO with exabyte scale and flexible pricing. Free migration tools handle the SQL translation automatically.
Try Free
Rate This Project
Login To Rate This Project

User Reviews

Be the first to post a review of PaddleSpeech!

Additional Project Details

Operating Systems

Mac, Windows

Programming Language

C++, Python

Related Categories

Python Voice Cloning Software, Python LLM Inference Tool, C++ Voice Cloning Software, C++ LLM Inference Tool

Registered

2023-03-23