Whisper
A widely deployed multilingual transcription and translation model.
Full Whisper record →Head to head
Whisper (OpenAI) and AssemblyAI (AssemblyAI) both sit in Audio. AssemblyAI shipped more tracked changes in the last 30 days (5 vs 2). Every value below comes from the latest crawl of the vendors' own pages.
| Attribute | Whisper OpenAI | AssemblyAI AssemblyAI |
|---|---|---|
| Segment | Audio | Audio |
| Version | gpt-transcribe | Universal-3.5 Pro |
| Entry cost | Pricing not specified in the document | Free tier with $50 free credits, then pay-as-you-go starting at $0.15/hr for pre-recorded STT (Universal-2) |
| Pricing tiers | Open Source Model free | Free Credits free |
| Model stack | gpt-transcribe · Whisper · Whisper large-v3 · whisper-1 | Claude · Claude 4.6 Sonnet · Claude 4.8 Opus · Gemini 2.5 Pro · Gemini 3.5 Flash · GPT-5.5 · Qwen3 Next 80B A3B · u3-rt-pro · Universal-2 · universal-3-pro · Universal-3.5 Pro · Universal-3.5 Pro Realtime · Universal-Streaming English · Universal-Streaming Multilingual |
| Context window | 30 second windows | n/a |
| Public API | ||
| Routes models | ||
| Changes / 30d | 2 | 5 |
| Origin | Global | Global |
| Capabilities | 99 languages, Translation to English, Word timestamps, Open weights, Robust to noise, File transcription, Language detection, Support for mp3, mp4, mpeg, mpga, m4a, wav, and webm formats | Speaker diarization, Summarisation, Topic detection, PII redaction, LeMUR framework, Pre-recorded STT, Real-time STT, Voice Agent API |
A widely deployed multilingual transcription and translation model.
Full Whisper record →Transcription plus summarisation, topic detection and speaker labels for developers.
Full AssemblyAI record →