PaddleSpeech is an open-source toolkit on PaddlePaddle platform for a variety of critical tasks in speech and audio, with state-of-art and influential models. Via the easy-to-use, efficient, flexible and scalable implementation, our vision is to empower both industrial application and academic research, including training, inference & testing modules, and deployment process. Low barriers to install, CLI, Server, and Streaming Server is available to quick-start your journey. We provide high-speed and ultra-lightweight models, and also cutting-edge technology. We provide production ready streaming asr and streaming tts system. Our frontend contains Text Normalization and Grapheme-to-Phoneme (G2P, including Polyphone and Tone Sandhi). Moreover, we use self-defined linguistic rules to adapt Chinese context.

Features

  • Implementation of critical audio tasks
  • Integration of mainstream models and datasets
  • Cascaded models application
  • Streaming ASR and TTS System
  • Rule-based Chinese frontend
  • The toolkit implements modules that participate in the whole pipeline of the speech tasks

Project Samples

Project Activity

See All Activity >

License

Apache License V2.0

Follow PaddleSpeech

PaddleSpeech Web Site

Other Useful Business Software
Train ML Models With SQL You Already Know Icon
Train ML Models With SQL You Already Know

BigQuery automates data prep, analysis, and predictions with built-in AI assistance.

Build and deploy ML models using familiar SQL. Automate data prep with built-in Gemini. Query 1 TB and store 10 GB free monthly.
Try Free
Rate This Project
Login To Rate This Project

User Reviews

Be the first to post a review of PaddleSpeech!

Additional Project Details

Operating Systems

Mac, Windows

Programming Language

C++, Python

Related Categories

Python Voice Cloning Software, Python LLM Inference Tool, C++ Voice Cloning Software, C++ LLM Inference Tool

Registered

2023-03-23