8 other companies on the FOBI voice & conversational interfaces roster in TTS / Voice Synthesis
ElevenLabs builds real-time text-to-speech and voice-cloning models with support for 29 languages and ultra-low latency.
Sesame builds fast, high-fidelity voice models for conversational AI with a focus on multilingual support and low-latency inference.
Cartesia develops ultra-low-latency voice models including Sonic, achieving ~40ms time-to-first-audio for real-time conversational AI.
Resemble AI provides voice cloning and custom text-to-speech APIs for gaming, entertainment, and call-center applications.
WellSaid Labs creates synthetic voice avatars for corporate training, e-learning, and media production with studio-quality output.
Murf AI provides a studio-quality text-to-speech platform with 120+ AI voices for video production, e-learning, and marketing content.
Rime develops fast, expressive text-to-speech models optimized for conversational AI and live-agent experiences.
Fish Audio builds multilingual TTS and voice-cloning models with support for Chinese, English, and other Asian languages.