Whisper

Whisper

OpenAI
+
+

Related Products

  • Google Cloud Speech-to-Text
    374 Ratings
    Visit Website
  • QEval
    30 Ratings
    Visit Website
  • Fathom
    7,272 Ratings
    Visit Website
  • Assembled
    233 Ratings
    Visit Website
  • CallTools
    494 Ratings
    Visit Website
  • Squaretalk
    258 Ratings
    Visit Website
  • DialerAI
    5 Ratings
    Visit Website
  • DialedIn
    577 Ratings
    Visit Website
  • Sendbird
    164 Ratings
    Visit Website
  • Docket
    58 Ratings
    Visit Website

About

IBM Watson® Speech to Text technology enables fast and accurate speech transcription in multiple languages for a variety of use cases, including but not limited to customer self-service, agent assistance and speech analytics. Get started fast with our advanced machine learning models out-of-the-box or customize them for your use case. Answer common call center queries using a Watson-powered virtual assistant on the phone. Improve call center performance by mining conversation logs to quickly and accurately identify emerging call patterns, customer complaints, sentiment, non-compliant behavior and more. Boost agent productivity and success with real time assistance during calls using AI-powered document and intranet search. As the agent is speaking with a customer, Watson listens in on the conversation, transcribes the audio, searches for relevant content within documentation, and feeds the answer back to the agent within seconds.

About

We’ve trained and are open-sourcing a neural net called Whisper that approaches human-level robustness and accuracy in English speech recognition. Whisper is an automatic speech recognition (ASR) system trained on 680,000 hours of multilingual and multitask supervised data collected from the web. We show that the use of such a large and diverse dataset leads to improved robustness to accents, background noise, and technical language. Moreover, it enables transcription in multiple languages, as well as translation from those languages into English. We are open-sourcing models and inference code to serve as a foundation for building useful applications and for further research on robust speech processing. The Whisper architecture is a simple end-to-end approach, implemented as an encoder-decoder Transformer. Input audio is split into 30-second chunks, converted into a log-Mel spectrogram, and then passed into an encoder.

Platforms Supported

Windows
Mac
Linux
Cloud
On-Premises
iPhone
iPad
Android
Chromebook

Platforms Supported

Windows
Mac
Linux
Cloud
On-Premises
iPhone
iPad
Android
Chromebook

Audience

Anyone in search of a solution to convert speech into text using AI-powered speech recognition

Audience

Anyone looking for a tool to recognize speech automatically and improve text transcription

Support

Phone Support
24/7 Live Support
Online

Support

Phone Support
24/7 Live Support
Online

API

Offers API

API

Offers API

Screenshots and Videos

Screenshots and Videos

Pricing

$0.01 per minute
Free Version
Free Trial

Pricing

No information available.
Free Version
Free Trial

Reviews/Ratings

Overall 0.0 / 5
ease 0.0 / 5
features 0.0 / 5
design 0.0 / 5
support 0.0 / 5

This software hasn't been reviewed yet. Be the first to provide a review:

Review this Software

Reviews/Ratings

Overall 0.0 / 5
ease 0.0 / 5
features 0.0 / 5
design 0.0 / 5
support 0.0 / 5

This software hasn't been reviewed yet. Be the first to provide a review:

Review this Software

Training

Documentation
Webinars
Live Online
In Person

Training

Documentation
Webinars
Live Online
In Person

Company Information

IBM
Founded: 1911
United States
www.ibm.com/cloud/watson-speech-to-text

Company Information

OpenAI
United States
openai.com/blog/whisper/

Alternatives

Alternatives

Azure AI Speech

Azure AI Speech

Microsoft
Transcribe

Transcribe

Wreally

Categories

Categories

Integrations

AI Sparks Studio
AnotherWrapper
Azure AI Speech
Baseten
Bolna
Hyprnote
IBM Cloud
IBM Cloud Pak for Applications
Kuku
MacWhisper
Monster API
OpenAI
ReByte
SheepScript.ai
Simplismart
Snippets AI
Tila
TurboScribe
Utterly Voice
VESSL AI

Integrations

AI Sparks Studio
AnotherWrapper
Azure AI Speech
Baseten
Bolna
Hyprnote
IBM Cloud
IBM Cloud Pak for Applications
Kuku
MacWhisper
Monster API
OpenAI
ReByte
SheepScript.ai
Simplismart
Snippets AI
Tila
TurboScribe
Utterly Voice
VESSL AI
Claim IBM Watson Speech to Text and update features and information
Claim IBM Watson Speech to Text and update features and information
Claim Whisper and update features and information
Claim Whisper and update features and information