Kunlun TechChina
Music generation with stem export and a public API.
- Version
- Mureka AI MV Director
- Cost
- Free tier · from $8/mo
- Model
- Mureka V8
OpenAI · #6 most active of 15 in Audio
A widely deployed multilingual transcription and translation model.
Trust score · Mixed
3 data points checked · 2026-09-11
Current version
gpt-transcribe
Entry cost
Pricing not specified in the document
Changes / 30d
2
The short answers · verified September 14, 2026
Source: OpenAI — openai.com · Reusable under CC BY 4.0 — cite Tomorrow
Open Source Model
Model weights and code available for free download
Free
Entry price over time
gpt-transcribe
OpenAI
Whisper
OpenAI
Whisper large-v3
OpenAI
whisper-1
OpenAI
Context window
30 second windows
Public API
yes
Multi-model routing
yes
Other tracked products running on the same foundation models — a quick read on how much of the catalog moves when one of these models changes.
green disclosed · amber inferred · grey unattributed. Aliases are folded into one model; provider concentration counts only models with an established vendor.
Reference data. Each figure is whatever the source actually said — weekly users, downloads, revenue run-rate — with its own definition and date. These are not comparable between tools and are never used to rank anything.
GitHub stars
106k
stars
as of Aug 2026·GitHub repository ·disclosed
GitHub stars
75k
stars
Developer-side proxy only.
as of Jun 2025·GitHub repository ·disclosed
Kunlun TechChina
Music generation with stem export and a public API.
Deepgram
Real-time speech to text
Suno
Full songs from a text prompt
AssemblyAI
Speech intelligence API
ElevenLabs
Lifelike text to speech
Cartesia
Ultra-low-latency voice models built on state space models.
New capabilities: audio transcription, speech to text
36 tracked38 tracked · +audio transcription, speech to text
sourceNew capabilities: Speech to text, Subtitles, Translations
33 tracked36 tracked · +Speech to text, Subtitles, Translations
sourceNew capabilities: Speech-to-text, Translation
26 tracked28 tracked · +Speech-to-text, Translation
sourceWhisper added Whisper to its model stack
Whisper large-v3, gpt-transcribe, whisper-1Whisper, Whisper large-v3, gpt-transcribe, whisper-1
sourceNew capabilities: Audio transcription, Audio translation, Support for mp3, mp4, mpeg, mpga, m4a, wav, webm
22 tracked25 tracked · +Audio transcription, Audio translation, Support for mp3, mp4, mpeg, mpga, m4a, wav, webm
sourceWhisper added whisper-1 to its model stack
Whisper large-v3, gpt-transcribeWhisper large-v3, gpt-transcribe, whisper-1
sourceWhisper added gpt-transcribe to its model stack
Whisper large-v3Whisper large-v3, gpt-transcribe
sourceNew capabilities: File transcription, Language detection, Support for mp3, mp4, mpeg, mpga, m4a, wav, and webm formats
55
sourceNew capabilities: speech recognition, multilingual transcription, zero-shot performance
29 tracked33 tracked · +speech recognition, multilingual transcription, zero-shot performance, timestamps
sourceStraight head-to-head pages against the busiest products in Audio.
Free to cite and reuse under CC BY 4.0. Permalink: https://tomorrow.aliensquad.ai/tools/whisper
Tomorrow. (2026). Whisper — version, pricing and model stack [Data set entry]. AlienSquad. Retrieved 2026-09-17, from https://tomorrow.aliensquad.ai/tools/whisper
@misc{tomorrow-tools-whisper,
author = {{Tomorrow}},
title = {Whisper — version, pricing and model stack},
year = {2026},
publisher = {AlienSquad},
howpublished = {\url{https://tomorrow.aliensquad.ai/tools/whisper}},
note = {Accessed: 2026-09-17}
}