Skip to content

Audio

Whisper

Open-source speech recognition model from OpenAI supporting 99 languages with state-of-the-art accuracy.

What’s good

  • Free and open source
  • Excellent accuracy
  • Many languages

Where it falls down

  • Requires technical setup
  • GPU recommended
  • No real-time

Features

Speech-to-text

99 languages

Translation

Local processing

Alternatives

Others doing the same job

All Audio tools →

ElevenLabs · Audio

Leading AI voice synthesis platform offering ultra-realistic text-to-speech, voice cloning, and dubbing.

Pricing

Free - $99/month

Trust 91

Visit

Speechmatics · Audio

An enterprise speech recognition engine known for accent inclusivity, supporting 50+ languages with real-time and batch transcription APIs.

Pricing

Free tier + usage-based

Trust 91

Visit

Sonix · Audio

An automated transcription, translation and subtitling platform with a polished editor, multi-language support and shareable media players.

Pricing

$22/month

Trust 90

Visit

Alitu · Audio

Alitu cleans, edits, and produces podcasts automatically with audio polish, music, and one-click publishing.

Pricing

$38-$48/month

Trust 88

Visit

Some links on this page are affiliate links, including links that pass through Skimlinks. If you buy through them we may earn a commission, at no extra cost to you. It never affects a tool's placement or score: see our methodology and affiliate disclosure. We currently earn nothing from this particular link.