Voice Generation & Conversion AI Tools
624 tools
624 tools · page 25 of 26
DupDub translates and dubs video content into other languages using AI voices while syncing lip movement, aimed at creators and studios localizing video for global audiences.
Maestra generates subtitles, dubbed audio, and transcripts for video content across many languages, aimed at content creators and media companies localizing video.
Free open-source AI voice typing app for macOS Windows and Linux
Real-time AI speech translation platform for video calls and live events
Papercup translates and dubs video content into other languages using AI voices reviewed by human linguists, aimed at media companies localizing content for global audiences.
Parloa builds AI voice agents for enterprise contact centers, with a strong presence in the European market, aimed at automating high-volume customer service phone calls.
PolyAI builds voice assistants that handle customer service phone calls for large enterprises like banks and retailers, aimed at replacing scripted IVR systems with natural conversation.
Regal.ai helps businesses trigger personalized AI-assisted phone calls and texts based on customer behavior, aimed at sales and retention teams running outbound engagement campaigns.
Retell AI provides infrastructure for building AI voice agents with an emphasis on low latency and voice quality, aimed at developers building inbound phone-based AI experiences.
Rev offers both AI-generated and human-reviewed transcription, captioning, and translation services, aimed at customers who need higher accuracy guarantees than AI alone provides.
Sonix transcribes audio and video into text and generates subtitles in many languages, aimed at podcasters, researchers, and video teams needing accurate searchable transcripts.
superwhisper runs locally on Mac to transcribe spoken words into text in any app, using AI to clean up grammar and formatting as the user speaks.
Synthflow lets non-technical teams build AI phone agents using a visual builder and templates, aimed at agencies and service businesses deploying voice AI without engineering resources.
AI voice platform for text-to-speech cloning and multilingual translation
Trint transcribes audio and video into searchable, editable text, with tools for collaborative editing and publishing aimed at newsrooms and media production teams.
AI voice generator and content creation tool with realistic AI voices
Vapi gives developers building blocks for creating custom voice AI agents, with flexibility to mix and match speech, language, and telephony providers for complex use cases.
Verbit provides AI-powered transcription, captioning, and live captioning services aimed at universities, courts, and enterprises with strict accuracy and compliance requirements.
Vocode is an open-source developer framework for building real-time voice AI agents, letting teams self-host and customize their voice stack rather than using a closed platform.
Whisper API provides a hosted transcription service built on OpenAI's Whisper speech recognition model, aimed at developers who want speech-to-text without running the model themselves.
Wispr Flow lets users dictate text anywhere on their computer using natural speech, with AI cleaning up filler words and formatting on the fly, aimed at replacing typing for everyday writing.
Amazon Polly converts text into lifelike speech across dozens of languages and voices through an API, aimed at developers adding narration or voice output to applications at cloud scale.