Whisper
Official
Model variants
Adds fast automatic speaker recognition with word-level timestamps and speaker diarization.
Faster reimplementation of Whisper using CTranslate2.
JAX implementation of Whisper for up to 70x speed-up on TPU.
Adds word-level timestamps and confidence scores.
Whisper running on OpenVINO.
Whisper running on TensorFlow Lite.
Various Whisper variants on Hugging Faces.
Whisper that can recognize non-speech audio events in addition to speech.
Apps
Audio transcription iOS and macOS app.
Audio transcription macOS app. (Freemium)
Audio transcription iOS app. (Freemium)
Audio journal iOS app.
Audio transcription macOS app.
Audio transcription and translation macOS app.
Audio transcription macOS app. (Freemium · Electron)
Audio/video management macOS app.
Global audio transcription macOS menu bar app.
Local speech-to-text transcription for macOS and Windows with system-wide dictation.
Audio transcription Linux app.
Dictation macOS app powered by OpenAI API.
Windows and macOS app for audio transcription and speaker diarization. (Freemium)
Real-time audio transcription on macOS and Windows. (Freemium · Electron)
Android app for transcription and translation. (FOSS)
Dictation and transcription macOS app. (FOSS)
AI voice dictation for Mac. (FOSS)
Dictation app for macOS. (FOSS)
24/7 local screen and audio recording with AI search. (FOSS)
Web apps
Hosted
Audio transcription and annotation tool.
Runs locally in your browser.
Transcription with real-time processing.
Local transcription using WebGPU, with optimised fine-tuned models for several languages. (FOSS)
Self-hosted
Subtitle generation.
GUI and API for Whisper.
Laravel app to transcribe and translate audio files.
Transcriptions, summary and more for meetings and any browser tab. (Chrome app)
CLI tools
YouTube subtitle generation.
Generate captions for videos.
Standalone Windows executable for Whisper and Faster Whisper.
Whisper command-line tool based on CTranslate2, compatible with the original.
Achieve transcription speeds near 30x real-time with several optimizations.
Automatic speech recognition with speaker diarization.
On-device speech-to-text CLI using faster-whisper with automatic clipboard copy.
Playgrounds
Whisper demo running on Hugging Faces. ([Source](https://huggingface.co/spaces/openai/whisper/tree/main))
Whisper demo running on Monster API. ([Source](https://github.com/saharmor/whisper-playground))
Whisper demo by Pluja. ([Source](https://codeberg.org/pluja/web-whisper))
Running on Colab.
Packages
JavaScript
React hook.
Articles
The future of machine learning lies in adaptable and accessible open-source speech-transcription programs.
Explains how to install and run the model, as well as providing a performance analysis comparing Whisper to other models.
The tutorial demonstrates Whisper's speech-to-text model, with a demo on running it in a Gradient Notebook and a guide for setting up a Flask app with Gradient Deployments.
Tutorial on the Whisper API with Python for speech-to-text transcription, showcasing GPU's faster transcription and advanced technology.
Videos
Introduction to Whisper.
Community
Third-party APIs
Extension of the Whisper model which adds powerful features such as speaker identification custom vocabulary, summarization, and chapter generation.
Use Whisper running on Replicate.
APIs that use Whisper.