Skip to content

Catalogaudio

WhisperX

radarBlueshift — gaining momentum

Open-source speech recognition toolkit with batched Whisper, word-level timestamps, and speaker diarization.

Visit github.com →Add to kitOpen in Noemium

ledger entry · observed by @catalog-radar

Open-source speech recognition toolkit with batched Whisper, word-level timestamps, and speaker diarization.

Radar · no verdict. Confirmed from primary sources; we have not issued ship, situational, or skip.

Price trailchangelog →

No price move in this repo yet.

Known limitations

  • Radar: not field-run in this catalog wave.
  • Overlapping speech is not handled well and speaker diarization is far from perfect, per the project's own limitations list.

Facts

pricing
free1
free tier
yes
open source
yes
api
no
self-host
yes
category
audio
evidence
radar
Models
bundled

Sources

  1. 1github.com/m-bain/whisperX[snapshot]

No source, no number.

Edit this page on GitHub

Similar joball alternatives →

toolverdictpriceossself-host
WhisperXradarfreeyesyes
Deepgramshipfreemiumnono
ElevenLabsshipfreemiumnono
Sunoshipfreemiumnono

Facts from the catalog files — not scores you can buy. Quality and speed stay in the briefing, not in a fake 1–5 grid.

Deepgram
audioship

Deepgram

Speech-to-text API built for developers — fast, cheap, real-time.

APIFREE TIER

freemium$200 free credits; Nova-3 STT from $0.0048/min (promo)

Steady momentum
ElevenLabs
audioship

ElevenLabsanchor

The reference standard for AI voice — TTS, cloning, dubbing and voice agents.

APIFREE TIER

freemiumfree tier (10k credits/mo, non-commercial); Starter $6/mo, Creator $22/mo

Blueshift — gaining momentum
Suno
audioship

Suno

Full-song generation — vocals, lyrics, arrangement — from a text prompt.

FREE TIER

freemiumfree tier (10 songs/day, non-commercial); Pro $10/mo, Premier $30/mo

Blueshift — gaining momentum