WhisperX: Automatic Speech Recognition with Word-level Timestamps (& Diarization)
WhisperX adds accurate word-level timestamps and speaker diarization on top of Whisper, ideal for subtitles and multi-speaker transcripts.
pip install whisperx

Excerpts from the project README on GitHub. Copyright and licensing remain with the respective authors.
Claim its page: indexed whatever its rank, translated into six languages, and enriched with what you write yourself.
Get an email alert on its next release or when it starts trending — never miss the moment.
Free · no card · unsubscribe anytimeWhisperX: Automatic Speech Recognition with Word-level Timestamps (& Diarization)
whisperX has 23.1k stars on GitHub. It has been forked 2.3k times. whisperX is written mainly in Python. It has been in active development since 2022. whisperX is available under the BSD-2-Clause license. Its main topics are asr, speech, speech-recognition, speech-to-text.
Read the full guideWhisperX: Automatic Speech Recognition with Word-level Timestamps (& Diarization)
whisperX is an open-source project. It is released under the BSD-2-Clause license.
Yes. whisperX is free and open source — you can use, modify and self-host it.
whisperX is available under the BSD-2-Clause license.
whisperX is written mainly in Python.
Add this live badge to your README — your GitHub stars and directory rank, refreshed daily.
[](https://olud.ai/project/m-bain-whisperx.html)
Measured from GitHub topics shared by both projects, weighted by how rare each topic is.