noemiumopen beta

Catalog/ № 063

Whisper

shipSteady momentum

OpenAI's open-weights speech recognition — the default transcription model.

github.comVerified 2026-08-15

ledger entry · observed by @whysanesanders

The reason paid transcription became a niche: large-v3 accuracy is production-grade across dozens of languages and it runs on a laptop. Every transcription pipeline starts here.

Known limitations

  • Hallucinates text on silence and noisy segments.
  • No built-in speaker diarization (pair with pyannote).
  • large-v3 needs a decent GPU for real-time work.
  • whisper-1 API is legacy; the current hosted line is the gpt-transcribe family.

Facts

pricing
free1
price note
open source (MIT), self-host free; hosted: gpt-4o-transcribe $0.006/min
free tier
yes
open source
yes
api
yes
self-host
yes
category
audio

Models used

whisper-v3

Sources

  1. 1github.com/openai/whisper

If we can't show the source, we don't print the number.

Edit this page on GitHub

Same shelf — audio