Whisper

Whisper

OpenAI
+
+

Related Products

  • Google Cloud Speech-to-Text
    375 Ratings
    Visit Website
  • QEval
    30 Ratings
    Visit Website
  • LM-Kit.NET
    25 Ratings
    Visit Website
  • Google AI Studio
    11 Ratings
    Visit Website
  • Enterprise Bot
    23 Ratings
    Visit Website
  • kama DEI
    8 Ratings
    Visit Website
  • Qminder
    337 Ratings
    Visit Website
  • Astra Pentest
    238 Ratings
    Visit Website
  • Soraban
    6 Ratings
    Visit Website
  • ClickLearn
    67 Ratings
    Visit Website

About

Intelligent Speech Interaction is developed based on state-of-the-art technologies such as speech recognition, speech synthesis, and natural language understanding. Enterprises can integrate Intelligent Speech Interaction into their products to enable them to listen, understand, and converse with users, providing users with an immersive human-computer interaction experience. Intelligent Speech Interaction is currently available in Mandarin Chinese, Cantonese Chinese, English, Japanese, Korean, French and Indonesian, and please stay tuned for other languages. Intelligent Speech Interaction is suitable for various scenarios, including intelligent Q&A, intelligent quality inspection, real-time subtitling for speeches, and transcription of audio recordings. Intelligent Speech Interaction has been successfully applied in many industries such as finance, insurance, eCommerce and smart home.

About

We’ve trained and are open-sourcing a neural net called Whisper that approaches human-level robustness and accuracy in English speech recognition. Whisper is an automatic speech recognition (ASR) system trained on 680,000 hours of multilingual and multitask supervised data collected from the web. We show that the use of such a large and diverse dataset leads to improved robustness to accents, background noise, and technical language. Moreover, it enables transcription in multiple languages, as well as translation from those languages into English. We are open-sourcing models and inference code to serve as a foundation for building useful applications and for further research on robust speech processing. The Whisper architecture is a simple end-to-end approach, implemented as an encoder-decoder Transformer. Input audio is split into 30-second chunks, converted into a log-Mel spectrogram, and then passed into an encoder.

Platforms Supported

Windows
Mac
Linux
Cloud
On-Premises
iPhone
iPad
Android
Chromebook

Platforms Supported

Windows
Mac
Linux
Cloud
On-Premises
iPhone
iPad
Android
Chromebook

Audience

Businesses seeking a speech recognition, speech synthesis, and natural language understanding solution

Audience

Anyone looking for a tool to recognize speech automatically and improve text transcription

Support

Phone Support
24/7 Live Support
Online

Support

Phone Support
24/7 Live Support
Online

API

Offers API

API

Offers API

Screenshots and Videos

Screenshots and Videos

Pricing

$1.40 per hour
Free Version
Free Trial

Pricing

No information available.
Free Version
Free Trial

Reviews/Ratings

Overall 0.0 / 5
ease 0.0 / 5
features 0.0 / 5
design 0.0 / 5
support 0.0 / 5

This software hasn't been reviewed yet. Be the first to provide a review:

Review this Software

Reviews/Ratings

Overall 0.0 / 5
ease 0.0 / 5
features 0.0 / 5
design 0.0 / 5
support 0.0 / 5

This software hasn't been reviewed yet. Be the first to provide a review:

Review this Software

Training

Documentation
Webinars
Live Online
In Person

Training

Documentation
Webinars
Live Online
In Person

Company Information

Alibaba Cloud
Founded: 2008
China
www.alibabacloud.com/product/intelligent-speech-interaction

Company Information

OpenAI
United States
openai.com/blog/whisper/

Alternatives

SpeechPulse

SpeechPulse

AV BEAM

Alternatives

Inworld TTS

Inworld TTS

Inworld
Transcribe

Transcribe

Wreally

Categories

Categories

Integrations

AI Sparks Studio
Baseten
Blink
Krater.ai
LastMile AI
LazyTyper
Monster API
OpenAI
Pruna AI
ReByte
SheepScript.ai
Shownotes
Simplismart
Thinkbuddy
Tila
TurboScribe
Undrstnd
Unremot
Zo

Integrations

AI Sparks Studio
Baseten
Blink
Krater.ai
LastMile AI
LazyTyper
Monster API
OpenAI
Pruna AI
ReByte
SheepScript.ai
Shownotes
Simplismart
Thinkbuddy
Tila
TurboScribe
Undrstnd
Unremot
Zo
Claim Alibaba Cloud Intelligent Speech Interaction and update features and information
Claim Alibaba Cloud Intelligent Speech Interaction and update features and information
Claim Whisper and update features and information
Claim Whisper and update features and information