Whisper

Whisper

OpenAI
+
+

Related Products

  • Google Cloud Speech-to-Text
    373 Ratings
    Visit Website
  • QEval
    30 Ratings
    Visit Website
  • Vertex AI
    783 Ratings
    Visit Website
  • Google AI Studio
    11 Ratings
    Visit Website
  • Fathom
    7,148 Ratings
    Visit Website
  • LALAL.AI
    4,565 Ratings
    Visit Website
  • Qloo
    23 Ratings
    Visit Website
  • LM-Kit.NET
    23 Ratings
    Visit Website
  • Docmosis
    48 Ratings
    Visit Website
  • CallHub
    424 Ratings
    Visit Website

About

Automatically convert audio and video files and live audio streams to text with AssemblyAI's speech-to-text APIs. Do more with audio intelligence, summarization, content moderation, topic detection, and more. Powered by cutting-edge AI models. From in-depth tutorials to detailed changelogs, to comprehensive documentation, AssemblyAI is focused on providing developers a great experience every step of the way. From core speech-to-text conversion to sentiment analysis, our simple API offers a full suite of solutions catered to all your business speech-to-text needs. We work with startups of all sizes, from early-stage startups to scale-ups, by providing cost-efficient speech-to-text solutions. We're built for scale. We process millions of audio files every day for hundreds of customers, including dozens of Fortune 500 enterprises. Universal-2: Our most advanced speech-to-text model captures the complexity of human speech for impeccable audio data that powers sharper insights.

About

We’ve trained and are open-sourcing a neural net called Whisper that approaches human-level robustness and accuracy in English speech recognition. Whisper is an automatic speech recognition (ASR) system trained on 680,000 hours of multilingual and multitask supervised data collected from the web. We show that the use of such a large and diverse dataset leads to improved robustness to accents, background noise, and technical language. Moreover, it enables transcription in multiple languages, as well as translation from those languages into English. We are open-sourcing models and inference code to serve as a foundation for building useful applications and for further research on robust speech processing. The Whisper architecture is a simple end-to-end approach, implemented as an encoder-decoder Transformer. Input audio is split into 30-second chunks, converted into a log-Mel spectrogram, and then passed into an encoder.

Platforms Supported

Windows
Mac
Linux
Cloud
On-Premises
iPhone
iPad
Android
Chromebook

Platforms Supported

Windows
Mac
Linux
Cloud
On-Premises
iPhone
iPad
Android
Chromebook

Audience

Companies requiring a solution to automatically convert audio, video files, and live audio streams to text

Audience

Anyone looking for a tool to recognize speech automatically and improve text transcription

Support

Phone Support
24/7 Live Support
Online

Support

Phone Support
24/7 Live Support
Online

API

Offers API

API

Offers API

Screenshots and Videos

Screenshots and Videos

Pricing

$0.00025 per second
Free Version
Free Trial

Pricing

No information available.
Free Version
Free Trial

Reviews/Ratings

Overall 0.0 / 5
ease 0.0 / 5
features 0.0 / 5
design 0.0 / 5
support 0.0 / 5

This software hasn't been reviewed yet. Be the first to provide a review:

Review this Software

Reviews/Ratings

Overall 0.0 / 5
ease 0.0 / 5
features 0.0 / 5
design 0.0 / 5
support 0.0 / 5

This software hasn't been reviewed yet. Be the first to provide a review:

Review this Software

Training

Documentation
Webinars
Live Online
In Person

Training

Documentation
Webinars
Live Online
In Person

Company Information

AssemblyAI
Founded: 2017
United States
www.assemblyai.com

Company Information

OpenAI
United States
openai.com/blog/whisper/

Alternatives

Alternatives

Transcribe

Transcribe

Wreally

Categories

Categories

Integrations

LazyTyper
Nekton.ai
Vocode
AnotherWrapper
Axis LMS
Fuser
Hyprnote
Krater.ai
NoteVocal
OpenAI
ReByte
Ruby
SheepScript.ai
Steamship
Tila
TurboScribe
TypeScript
Undrstnd
brancher.ai

Integrations

LazyTyper
Nekton.ai
Vocode
AnotherWrapper
Axis LMS
Fuser
Hyprnote
Krater.ai
NoteVocal
OpenAI
ReByte
Ruby
SheepScript.ai
Steamship
Tila
TurboScribe
TypeScript
Undrstnd
brancher.ai
Claim AssemblyAI and update features and information
Claim AssemblyAI and update features and information
Claim Whisper and update features and information
Claim Whisper and update features and information