PaddleSpeech is an open-source toolkit on PaddlePaddle platform for a variety of critical tasks in speech and audio, with state-of-art and influential models. Via the easy-to-use, efficient, flexible and scalable implementation, our vision is to empower both industrial application and academic research, including training, inference & testing modules, and deployment process. Low barriers to install, CLI, Server, and Streaming Server is available to quick-start your journey. We provide high-speed and ultra-lightweight models, and also cutting-edge technology. We provide production ready streaming asr and streaming tts system. Our frontend contains Text Normalization and Grapheme-to-Phoneme (G2P, including Polyphone and Tone Sandhi). Moreover, we use self-defined linguistic rules to adapt Chinese context.

Features

  • Implementation of critical audio tasks
  • Integration of mainstream models and datasets
  • Cascaded models application
  • Streaming ASR and TTS System
  • Rule-based Chinese frontend
  • The toolkit implements modules that participate in the whole pipeline of the speech tasks

Project Samples

Project Activity

See All Activity >

License

Apache License V2.0

Follow PaddleSpeech

PaddleSpeech Web Site

Other Useful Business Software
Try Google Cloud Risk-Free With $300 in Credit Icon
Try Google Cloud Risk-Free With $300 in Credit

No hidden charges. No surprise bills. Cancel anytime.

Use your credit across every product. Compute, storage, AI, analytics. When it runs out, 20+ products stay free. You only pay when you choose to.
Start Free
Rate This Project
Login To Rate This Project

User Reviews

Be the first to post a review of PaddleSpeech!

Additional Project Details

Operating Systems

Mac, Windows

Programming Language

C++, Python

Related Categories

Python Voice Cloning Software, Python LLM Inference Tool, C++ Voice Cloning Software, C++ LLM Inference Tool

Registered

2023-03-23