openai/whisper
This system is an audio transcription service built on the Whisper model, responsible for converting speech to text with word-level timestamps. It handles audio processing, text normalization, and tokenization using the tiktoken library. The codebase includes comprehensive tests for timing, tokenization, and transcription accuracy.
63.4
Adequate · 5 August 2026
3.4k
lines of production code
Python
primary language
2
bus factor · 85 authors in all
4
measurements over time
Survey your own repository
openai/whisper was measured the same way every project in this corpus was: the same rubric, at a pinned commit, with the result published in full. Point the surveyor at a repository you know and see whether you agree with it.
About this page
- The description of this project is derived from its own commit history, not from its README.
- The score is its highest published measurement, taken on 5 August 2026 at a pinned commit. It is not a live figure and does not change until the project is measured again.
- Measured at commit 5f86d1d863 — the exact code this score is about.
- Scored under rubric rubric-2026.08.19. Score the same commit under that rubric and you get the same number.