Free Voice & Speech AI Tools
Discover 2065+ free and freemium AI tools across 50+ categories. Filter by category, use case, and features. Updated daily.
NobodyWho
NobodyWho is an open-source, on-device inference library and SDK that runs text, vision, and speech models (LLM, STT, TTS) locally across platforms, emphasizing performance, privacy, and offline use.
Voice & Speech
precallai
PreCallAI is a developer-first, production voice agent platform that orchestrates streaming STT → LLM → TTS call pipelines while letting customers bring their own provider keys and pay flat monthly plans (no per-minute billing).
Voice & Speech
callin-io
Callin.io is an enterprise-focused white-label AI voice platform that provides ultra-low latency voice AI (Neuron 1.0) and autonomous AI agents for industry-specific conversational automation, built for rapid deployment and enterprise security.
Voice & Speechsynthflow-ai
Synthflow is an enterprise-grade, end-to-end Voice AI platform that designs, deploys, and operates AI voice agents for inbound and outbound phone automation, combining a visual flow builder, in-house telephony, testing and monitoring, and deep integrations for B2B use cases.
Voice & Speech
bland-ai
Bland is an enterprise voice AI platform that builds, runs, and monitors production phone agents for regulated industries (healthcare, insurance, financial services, logistics), running on customer-controlled infrastructure and designed for high-security, high-stakes phone calls.
Voice & Speech
Relyable
Relyable is a simulation and monitoring platform for AI voice agents that generates realistic test conversations, evaluates every call against customizable rubrics, and monitors live production calls with real-time alerts to help teams ship high-performing voice agents faster.
Voice & Speech
Vogent Voicelab
Vogent Voicelab is a public-beta, high-quality text-to-speech platform and API that hosts state-of-the-art voice models (e.g., Sesame CSM-1B, Dia, Chatterbox, Orpheus), offering ultra-fast real-time inference, zero-shot voice cloning, hosted fine-tuning, and scalable deployment (including on-prem/VPC) for developers and enterprises.
Voice & Speech
Deepgram
Deepgram provides enterprise Voice AI APIs — real-time and batch speech-to-text (STT), text-to-speech (TTS), audio intelligence, and unified Voice Agent capabilities (including LLM orchestration) for developers, platforms, and enterprises.
Voice & Speech
Murf AI
Murf AI is an end-to-end AI voice platform providing production-grade text-to-speech APIs (Murf Falcon), a Studio for fast voiceover creation, and AI dubbing and agent solutions to build scalable conversational voice agents and localized audio content.
Voice & SpeechSpeechify
Speechify is a cross-platform Voice AI productivity assistant that provides text-to-speech, AI voice generation, voice typing (dictation), voice cloning, dubbing, and developer APIs to read, write, summarize, and create audio content across web, desktop, and mobile devices.
Voice & SpeechELSA Speak
ELSA Speak is an AI-powered English speaking coach (mobile app and B2B/Schools offerings) that provides personalized lessons, real-time pronunciation feedback, role-play conversations, and progress tracking to help learners improve fluency and confidence.
Voice & Speech