ElevenLabs is an AI audio platform that provides lifelike text-to-speech, voice cloning, dubbing, and sound effects generation for content creators and enterprises.
Add speech to your product
Transcription, understanding, and lifelike voices behind one API call, accurate enough to build on.
Cartesia builds voice model infrastructure focused on real-time performance for conversational interfaces and AI voice products.
Deepgram provides speech-to-text and voice understanding APIs for teams building transcription, voice workflows, and AI-powered audio products.
AssemblyAI provides speech-to-text, real-time transcription, and voice intelligence models for teams shipping voice AI and audio-driven workflows.
Open-source non-autoregressive zero-shot TTS model with flow matching for natural speech synthesis.
Cloud communications platform with AI-powered voice bots and natural language processing.
AI music and vocal generation platform that creates complete songs with vocals, instruments, and lyrics from text prompts.
PlayHT offers text-to-speech and voice generation infrastructure for products, media workflows, and conversational AI applications.
Resemble AI provides real-time AI voice generation, voice cloning, and speech-to-speech conversion with deepfake detection capabilities.
Open-source voice synthesis platform with high-quality voice cloning and multilingual text-to-speech generation.
AI media platform with speech recognition, content indexing, and voice analytics for media and government.
AI audio research platform with text-to-audio models Bark and Chirp for speech and music generation.
Real-time AI voice changer with celebrity and custom voices for streamers, gamers, and content creators.
Text-to-speech app converting any text into natural-sounding audio for reading documents, articles, and books.
AI text-to-speech platform with natural voices for reading documents, websites, and ebooks.
Real-time speech understanding and accent translation platform for contact centers.
Enterprise text-to-speech platform providing AI voices for e-learning, accessibility, and customer experience.
AI voiceover and text-to-speech platform with 500+ voices in 100 languages for video and content creation.
Voice AI platform powering conversational interfaces for automotive, IoT, and customer service applications.
Speech recognition and clinical documentation platform now from Solventum enabling voice-driven EHR workflows.
Human and AI transcription service providing accurate captions, subtitles, and transcripts for audio and video.
AI music generation platform creating professional-quality songs with realistic vocals across multiple genres and styles.
Chinese speech AI platform offering voice services, translation, and enterprise AI solutions.
Tactiq provides real-time meeting transcription and AI summaries for video calls.
Vapi is a developer platform for building, testing, and deploying voice AI agents that can handle phone calls with natural conversation capabilities and tool integrations.
AI audio transcription API with enterprise-grade speech-to-text, translation, and audio intelligence features.
Ultra-low-latency voice AI and text-to-speech platform for developers.
Speech-to-text API by Rev providing accurate automatic transcription with speaker diarization and custom vocabulary.
Voice interface API for building real-time spoken language understanding into applications and devices.
Real-time speech recognition platform with state-of-the-art accuracy for transcription and voice applications.