Skip to content
CAI
Produce a survey ↗Verify a survey

openai/whisper

This system is an audio transcription service built on the Whisper model, responsible for converting speech to text with word-level timestamps. It handles audio processing, text normalization, and tokenization using the tiktoken library. The codebase includes comprehensive tests for timing, tokenization, and transcription accuracy.

63.4

Adequate · 5 August 2026

3.4k

lines of production code

Python

primary language

2

bus factor · 85 authors in all

4

measurements over time

CAI band scale
CAI trend line
CAI lens gauges

Survey your own repository

openai/whisper was measured the same way every project in this corpus was: the same rubric, at a pinned commit, with the result published in full. Point the surveyor at a repository you know and see whether you agree with it.

Survey a repository

About this page

  • The description of this project is derived from its own commit history, not from its README.
  • The score is its highest published measurement, taken on 5 August 2026 at a pinned commit. It is not a live figure and does not change until the project is measured again.
  • Measured at commit 5f86d1d863 — the exact code this score is about.
  • Scored under rubric rubric-2026.08.19. Score the same commit under that rubric and you get the same number.
CAI link cards