Voice Generation & Conversion AI Tools
624 tools
624 tools · page 26 of 26
Azure AI Speech provides text-to-speech, speech-to-text, and voice translation services through Microsoft's cloud platform, aimed at enterprises building voice features into large-scale applications.
Cartesia builds the Sonic family of text-to-speech models optimized for very low latency, aimed at developers building voice agents and apps that need near-instant, emotionally expressive speech.
Google Cloud Text-to-Speech converts text into speech using WaveNet and Chirp voice models across many languages, aimed at developers building voice features into apps and IVR systems.
Hume AI builds an empathic voice interface that generates speech with nuanced emotional tone and can also analyze emotional expression in a user's voice, aimed at developers building more natural-sounding voice assistants.
Kits.ai lets musicians convert a vocal recording into a different AI-modeled singing voice, licensed with artist consent, aimed at producers experimenting with vocal styles for original music.
Listnr generates natural-sounding voiceovers from text across many languages and voices, aimed at podcasters and video creators who need narration without recording it themselves.
LOVO generates voiceovers in many languages and accents and offers dubbing tools for localizing videos, aimed at content teams producing multilingual media.
Murf AI creates studio-quality voiceovers from text, aimed at marketing videos, e-learning, and presentations, with a library of voices and an editor for timing and emphasis.
Musicfy lets users create AI voice covers, isolate vocals and instruments, and generate original songs from prompts, aimed at musicians and fans experimenting with AI-assisted covers.
NaturalReader converts documents, PDFs, and web pages into spoken audio using natural-sounding voices, aimed at students and professionals who prefer listening to dense reading material.
Otter.ai transcribes meetings in real time, generates summaries and action items, and integrates with video-call platforms like Zoom and Google Meet.
Play.ht converts text into speech using a large library of AI voices, offering both a web app and an API for embedding voice generation into other products.
PlayAI lets businesses build voice agents that can hold natural phone or app conversations, built by the team behind Play.ht's text-to-speech technology, aimed at customer service and outbound calling use cases.
Replica Studios provides licensed AI voice actors for video games, generating dialogue and character voices, aimed at game studios that need scalable voiceover without booking every line with human actors.
Resemble AI clones a person's voice from sample audio and generates new speech in that voice, with real-time voice conversion tools aimed at developers and content teams.
Respeecher converts one actor's voice performance into another target voice while preserving emotion and timing, used professionally in film dubbing and video game voice production.
Speechify reads articles, PDFs, and books aloud using natural-sounding AI voices, aimed at people who prefer listening to text or have reading difficulties like dyslexia.
Synthesys generates voiceovers and talking-avatar videos from text scripts, aimed at marketers and course creators producing videos without filming.
TTSMaker converts typed text into downloadable speech audio for free, across a range of languages and voices, aimed at users who need occasional narration without signing up for a subscription tool.
Voice-Swap lets producers apply licensed AI artist voice models to vocal tracks, with royalties routed back to the voice's rights holder, aimed at legally navigating AI voice covers in music.
Voice.ai transforms a user's voice into different character voices in real time during gaming, streaming, or calls, aimed at gamers and streamers wanting a lighter alternative to Voicemod.
Voicemaker converts text into speech across a large catalog of AI voices and languages, aimed at creators needing voiceover for videos, e-learning, and IVR systems on a budget.
Voicemod alters a user's voice in real time with different effects and character voices, popular among gamers and streamers for voice chat and content creation.
WellSaid Labs produces high-fidelity synthetic voices for corporate training, marketing, and product narration, aimed at teams needing professional broadcast-quality voiceover.