AssemblyAI
Speech intelligence API
- Version
- Universal-2
- Cost
- Pay as you go from $0.12/hr
- Model
- Universal-2
Voice, music, transcription and dubbing
15 tools · 1 changes / 30d
AssemblyAI
Speech intelligence API
Cartesia
Ultra-low-latency voice models built on state space models.
Deepgram
Real-time speech to text
ElevenLabs
Lifelike text to speech
Fireflies.ai
Notetaker with CRM-aware conversation intelligence.
Fish AudioChina
Open-weight multilingual TTS and voice cloning.
Granola
Local-first meeting notes that augment your own typing.
MiniMaxChina
Multilingual TTS and voice cloning at low per-character rates.
Kunlun TechChina
Music generation with stem export and a public API.
Otter.ai
Meeting transcription and AI meeting agents.
PlayAI
Voice agents and TTS for conversational products.
Speechmatics
Accuracy-focused speech recognition for broadcast and enterprise.
Suno
Full songs from a text prompt
Udio
Music generation and remixing
OpenAI
Open-weight speech recognition
MiniMax Speech — Emotion control exposed in the speech API
Voice cloning, 30+ languagesVoice cloning, 30+ languages, Emotion control
sourceElevenLabs — Music generation added to the platform
Voice cloning, Dubbing studioVoice cloning, Dubbing studio, Music
source