Audio

AssemblyAI

AssemblyAI · #2 most active of 15 in Audio

Compare →

Transcription plus summarisation, topic detection and speaker labels for developers.

Current version

Universal-2

Entry cost

Pay as you go from $0.12/hr

Changes / 30d

0

Capabilities

  • Speaker diarization
  • Summarisation
  • Topic detection
  • PII redaction
  • LeMUR framework

Pricing

  • Free credits

    Free

  • Pay as you go

    $0.12 per hour

    Free

  • Enterprise

    custom

    Free

Entry price over time

Not enough pricing history yet — we start charting from the second observation.

Geek mode

Models under the hood

  • Universal-2

    AssemblyAI

    disclosed
  • Claude

    Anthropic

    disclosed

Context window

n/a

Public API

yes

Multi-model routing

yes

Stack signals

  • REST API
  • Streaming WebSocket
  • SDKs

Shared model stack

Other tracked products running on the same foundation models — a quick read on how much of the catalog moves when one of these models changes.

green disclosed · amber inferred · grey unknown

Reported scale

Reference data. Each figure is whatever the source actually said — weekly users, downloads, revenue run-rate — with its own definition and date. These are not comparable between tools and are never used to rank anything.

No public usage figure on record for this product.

Also in Audio

All Audio
MiniMax Speech

MiniMaxChina

1

Multilingual TTS and voice cloning at low per-character rates.

Version
2.6
Cost
Free tier, API per character
Model
MiniMax Speech 2.6
Cartesia

Cartesia

Ultra-low-latency voice models built on state space models.

Version
Sonic 3
Cost
Free tier, $5/mo
Model
Sonic 3
Deepgram

Deepgram

Real-time speech to text

Version
Nova-3
Cost
Pay as you go from $0.0043/min
Model
Nova-3
ElevenLabs

ElevenLabs

Lifelike text to speech

Version
v3
Cost
Free tier, Starter $5/mo
Model
Eleven v3
Fireflies.ai

Fireflies.ai

Notetaker with CRM-aware conversation intelligence.

Version
2025
Cost
Free tier, $18/mo Pro
Model
in-house ASR
Fish Audio

Fish AudioChina

Open-weight multilingual TTS and voice cloning.

Version
S1
Cost
Open weights; API from $15/M chars
Model
Fish Speech S1

Elsewhere on the same models

Products in other segments that name one of AssemblyAI's models in their stack.

Enterprise assistant grounded in a company's AWS-connected content.

Version
2025
Cost
from $3/mo
Model
Claude

Search, chat and agents across Jira, Confluence and connected SaaS.

Version
2025
Cost
Included with paid Jira/Confluence plans
Model
Claude
Decagon

Decagon

AI support agents for high-volume consumer brands.

Version
2025
Cost
Usage-based, enterprise
Model
Claude
Devin

Cognition

Autonomous software engineer that works asynchronous tickets.

Version
Devin 2.x
Cost
from $20/mo
Model
Claude
Elicit

Elicit

AI research assistant for papers

Version
2025 platform
Cost
Free tier, Plus $12/mo
Model
Claude
1

Prompt-to-design, Make and AI editing inside the design tool.

Version
2025
Cost
Included in paid seats with AI credits
Model
Claude

Change history

  • version

    Universal-2 model rolled out as the default

    Universal-1Universal-2

    source