Audio

MiniMax Speech

MiniMax · #11 most active of 15 in Audio

Compare →

MiniMax's speech stack covers 30+ languages with fast cloning from short samples.

100

Trust score · Verified

2 data points checked · 2026-09-14

Model stack: VerifiedEntry price: UnverifiableVersion: VerifiedHow this works →

Current version

2.8

Entry cost

from $30/mo

Changes / 30d

0

The short answers

What is the latest version of MiniMax Speech?
The current shipped version of MiniMax Speech is 2.8.
How much does MiniMax Speech cost?
MiniMax Speech starts at $30 per month on its cheapest paid tier.
What AI model does MiniMax Speech use?
MiniMax Speech runs primarily on MiniMax Music 3.0.

Source: MiniMaxwww.minimax.io · Reusable under CC BY 4.0 — cite Tomorrow

Capabilities

  • Voice cloning
  • 30+ languages
  • Emotion control
  • Streaming
  • Speech Generation
  • Audio Generation
  • API
  • speech generation
  • audio generation
  • Speech Synthesis
  • Speech generation
  • Audio generation
  • API Token Plan
  • Audio
  • music generation
  • text-to-speech
  • voice-generation
  • Text-to-Speech
  • Speech
  • audio
  • api
  • speech synthesis

Pricing

  • Free

    Free

  • API per 1M chars

    $30

Entry price over time

7/22/2026 · $309/14/2026 · $30

Geek mode

Models under the hood

  • MiniMax Music 3.0

    MiniMax

    disclosed
  • MiniMax Speech 2.8

    MiniMax

    disclosed
  • Speech 2.8

    MiniMax

    disclosed

Context window

n/a

Public API

yes

Multi-model routing

yes

Stack signals

  • Streaming API
  • Voice library

Shared model stack

Other tracked products running on the same foundation models — a quick read on how much of the catalog moves when one of these models changes.

green disclosed · amber inferred · grey unattributed. Aliases are folded into one model; provider concentration counts only models with an established vendor.

Reported scale

Reference data. Each figure is whatever the source actually said — weekly users, downloads, revenue run-rate — with its own definition and date. These are not comparable between tools and are never used to rank anything.

No public usage figure on record for this product.

Also in Audio

All Audio
Mureka

Kunlun TechChina

7

Music generation with stem export and a public API.

Version
Mureka AI MV Director
Cost
Free tier · from $8/mo
Model
Mureka V8
Deepgram

Deepgram

6

Real-time speech to text

Version
Nova-3
Cost
Free tier · from $4,000/min
Model
Aura-1
Suno

Suno

6

Full songs from a text prompt

Version
v6
Cost
Free tier · from $10/mo
Model
Bark
AssemblyAI

AssemblyAI

5

Speech intelligence API

Version
Universal-3.5 Pro
Cost
Free tier with $50 free credits, then pay-as-you-go starting at $0.15/hr for pre-recorded STT (Universal-2)
Model
Claude
ElevenLabs

ElevenLabs

3

Lifelike text to speech

Version
Eleven v3
Cost
Free tier · from $6/mo
Model
Dubbing v2
Whisper

OpenAI

2

Open-weight speech recognition

Version
gpt-transcribe
Cost
Pricing not specified in the document
Model
gpt-transcribe

Change history

  • capability

    New capabilities: Speech

    18 tracked19 tracked · +Speech

    source
  • version

    MiniMax Speech moved to 2.8

    Speech 2.82.8

    source
  • capability

    New capabilities: text-to-speech, voice-generation

    15 tracked17 tracked · +text-to-speech, voice-generation

    source
  • model

    QA auto-correction: model stack updated to MiniMax Speech 2.8, MiniMax Music 3.0 after a unanimous jury read of the source page

    MiniMax Speech 2.6, MiniMax Speech 2.8, Music 3.0, Speech 2.8MiniMax Speech 2.8, MiniMax Music 3.0

    source
  • capability

    New capabilities: API Token Plan

    123

    source
  • capability

    New capabilities: Speech generation, Audio generation

    102

    source
  • capability

    New capabilities: Speech Synthesis

    93

    source
  • capability

    New capabilities: speech generation, audio generation

    72

    source
  • version

    MiniMax Speech moved to 2.8

    2.62.8

    source
  • model

    MiniMax Speech added MiniMax Speech 2.8 to its model stack

    MiniMax Speech 2.6MiniMax Speech 2.6, MiniMax Speech 2.8

    source
  • capability

    New capabilities: Speech Generation, Audio Generation, API

    43

    source
  • capability

    Emotion control exposed in the speech API

    Voice cloning, 30+ languagesVoice cloning, 30+ languages, Emotion control

    source
  • version

    MiniMax Speech moved to Speech 2.8

    2.8Speech 2.8

    source
  • model

    MiniMax Speech added Speech 2.8 to its model stack

    MiniMax Speech 2.6, MiniMax Speech 2.8MiniMax Speech 2.6, MiniMax Speech 2.8, Speech 2.8

    source
  • version

    MiniMax Speech moved to 2.8

    Speech 2.82.8

    source
  • capability

    New capabilities: Audio

    13 tracked14 tracked · +Audio

    source
  • model

    MiniMax Speech added Music 3.0 to its model stack

    MiniMax Speech 2.6, MiniMax Speech 2.8, Speech 2.8MiniMax Speech 2.6, MiniMax Speech 2.8, Music 3.0, Speech 2.8

    source
  • capability

    New capabilities: music generation

    14 tracked15 tracked · +music generation

    source
  • model

    MiniMax Speech added Speech 2.8 to its model stack

    MiniMax Music 3.0, MiniMax Speech 2.8MiniMax Music 3.0, MiniMax Speech 2.8, Speech 2.8

    source
  • version

    MiniMax Speech moved to Speech 2.8

    2.8Speech 2.8

    source
  • version

    MiniMax Speech moved to 2.8

    Speech 2.82.8

    source
  • version

    MiniMax Speech moved to Speech 2.8

    2.8Speech 2.8

    source
  • capability

    New capabilities: Text-to-Speech

    17 tracked18 tracked · +Text-to-Speech

    source
  • version

    MiniMax Speech moved to 2.8

    Speech 2.82.8

    source
  • version

    MiniMax Speech moved to 2.8

    Speech 2.82.8

    source
  • capability

    New capabilities: audio

    19 tracked20 tracked · +audio

    source
  • version

    MiniMax Speech moved to Speech 2.8

    2.8Speech 2.8

    source
  • capability

    New capabilities: api

    20 tracked21 tracked · +api

    source
  • version

    MiniMax Speech moved to Speech 2.8

    2.8Speech 2.8

    source
  • capability

    New capabilities: speech synthesis

    21 tracked22 tracked · +speech synthesis

    source
  • version

    MiniMax Speech moved to 2.6

    Speech 2.52.6

    source
  • model

    MiniMax Speech added MiniMax Speech 2.6 to its model stack

    MiniMax Speech 02, MiniMax Speech 2.5MiniMax Speech 02, MiniMax Speech 2.5, MiniMax Speech 2.6

    source
  • capability

    New capabilities: Text-to-Speech, Speech-to-Text

    3 tracked5 tracked · +Text-to-Speech, Speech-to-Text

    source
  • version

    MiniMax Speech moved to Speech 2.5

    2.5Speech 2.5

    source
  • model

    MiniMax Speech added MiniMax Speech 02, MiniMax Speech 2.5 to its model stack

    Speech 02, Speech 2.5MiniMax Speech 02, MiniMax Speech 2.5, Speech 02, Speech 2.5

    source
  • capability

    New capabilities: Text-to-speech, Speech-to-text, Audio generation

    3 tracked6 tracked · +Text-to-speech, Speech-to-text, Audio generation

    source
  • version

    MiniMax Speech moved to 2.5

    Speech-02-hd2.5

    source
  • model

    MiniMax Speech added Speech 02, Speech 2.5 to its model stack

    Speech-02-hdSpeech 02, Speech 2.5, Speech-02-hd

    source
  • capability

    New capabilities: Audio Generation, API

    3 tracked5 tracked · +Audio Generation, API

    source
  • version

    MiniMax Speech moved to Speech-02-hd

    Speech-02-hd

    source
Subscribe to MiniMax Speech changes

MiniMax Speech compared

Straight head-to-head pages against the busiest products in Audio.

How to cite this page

Free to cite and reuse under CC BY 4.0. Permalink: https://tomorrow.aliensquad.ai/tools/minimax-speech

APA
Tomorrow. (2026). MiniMax Speech — version, pricing and model stack [Data set entry]. AlienSquad. Retrieved 2026-09-17, from https://tomorrow.aliensquad.ai/tools/minimax-speech
BibTeX
@misc{tomorrow-tools-minimax-speech,
  author       = {{Tomorrow}},
  title        = {MiniMax Speech — version, pricing and model stack},
  year         = {2026},
  publisher    = {AlienSquad},
  howpublished = {\url{https://tomorrow.aliensquad.ai/tools/minimax-speech}},
  note         = {Accessed: 2026-09-17}
}