Advanced AI text-to-speech (TTS) and voice synthesis platform.
22 subcategories · 2,095 tools · Updated September 29, 2026
Voice Generation & Conversion covers AI tools that move between speech and text or change how a voice sounds. In one direction, AI Transcription and AI Speech-to-Text tools turn recordings, calls and video into transcripts with timestamps, speaker labels and subtitles. In the other, AI Text-to-Speech and voice-over tools read scripts aloud in synthetic voices you can choose by language, accent and style. Between the two sit voice cloning, voice changers and AI Dubbing tools, which recreate a specific voice or carry speech into another language. The group also includes voice assistants you can talk to, podcast tools for editing and clipping episodes, and enhancers that clean up noisy recordings.
For transcription, test accuracy on your own audio, since accents, crosstalk and specialist vocabulary are where tools differ, and check language support, speaker identification and export formats such as SRT or DOCX. For voice generation, listen to long passages rather than short demos, check pronunciation controls, and confirm whether commercial use and monetized platforms are allowed on your plan. Pricing is usually per minute of audio or per character, so estimate your monthly volume. With cloning, only use voices you have permission to use; many services require consent verification and restrict imitating public figures.
3,888 listings · open a subcategory to compare every tool in it
Voice chat tools let you talk with AI characters and assistants by speech or text.
Upload a recording or video, or paste a link, and get the spoken content back as text: full transcripts, timed subtitles…
Text-to-speech tools use AI voices to read your scripts aloud and return audio you can download.
Voice assistants combine speech recognition with language understanding so people can talk to software and have it act.
Speech recognition tools listen to recorded or live audio and turn it into text you can read, search and edit.
Speech synthesis tools take written text or a sample recording and produce spoken audio.
Recording tools capture what happens on your screen, in your microphone, or during a video call, then help you do something with the result.
These tools generate speech that imitates the voice of a celebrity, public figure, or a well-known anime, game or cartoon character.
Voice cloning tools recreate a specific voice from a short recording, then read any script in that voice.
Voice generators for short-form video turn a written script into spoken narration you can drop straight into TikTok clips.
Podcast editing tools use AI for the audio work that follows recording: cleaning up noise, cutting filler words and long silences, balancing levels…
AI dubbing tools take an existing video or audio file and produce a spoken track in another language.
Turn articles, documents, links, and notes into podcast-style audio with synthetic hosts, multiple languages, and voice options.
24 picks from 2,095 tools, each listed once. Ranked by traffic to the tool's own website, GitHub stars, subcategory coverage and user saves.
Advanced AI text-to-speech (TTS) and voice synthesis platform.
Adobe Podcast offers advanced AI-powered audio recording and editing directly from the web.
AI-powered tool for transcribing audio and video files to text effortlessly.
Luvvoice is a free text-to-speech tool supporting over 70 languages and 200 voices.
Transform your audio with Fish Audio's innovative tools.
A free online tool to transform and modify voices with various effects.
FineVoice is a versatile AI voice generator. Instantly create high-quality, royalty-free voices, SFX, and music.
Speechify converts text into natural-sounding speech across a variety of platforms.
Typecast offers advanced AI voice generation and virtual avatars for engaging content creation.
Rev AI provides automated transcription and captioning services powered by advanced AI technology.
NaturalReader converts text to natural-sounding speech.
KikiVoice delivers realistic AI text-to-speech and voice cloning for creators, podcasts, and interactive content.
Experience truly uncensored AI with limitless capabilities.
AnySpeech is an AI voice studio that turns text into natural speech with 100+ voices in 50+ languages.
FakeYou offers AI-powered text-to-speech and voice generation services.
An AI-powered tool for editing videos and podcasts with ease.
Create engaging videos with EchoWave's online video and audio editor.
Captions is an AI-powered creative studio for creating studio-grade videos effortlessly.
Kits AI offers advanced AI tools for music production and voice transformation.
AI Dubbing provides natural, high-quality video dubbing with support for over 20 languages and 100+ voice tones.
Panda Video is a versatile video hosting platform designed for infoproducers.
Tarteel is an AI-powered app for Quran recitation and memorization.
AI-powered app for oral language skills improvement.
Tapesearch is a search engine for podcast transcripts.
The terms overlap. AI Transcription tools are built around files and finished results: you upload a recording and get an edited transcript, often with speaker names, summaries and subtitle export. AI Speech Recognition tools focus on the underlying engine, frequently live or through an API, for developers who want to add voice input to an app. If you just need readable text from meetings or interviews, start with transcription tools.
Cloning your own voice, or a voice you have written permission to use, is generally fine. Cloning someone else without consent can break platform rules and, in many places, personality, privacy or fraud laws, especially if the result is used to deceive or to advertise. Reputable AI Voice Cloning services ask you to verify consent and label synthetic audio. Check the terms before using cloned or celebrity-style voices in anything public.
Listen to several minutes of your own script, not just a demo sentence, and pay attention to names, numbers and pauses. Check whether you can adjust speed, emphasis and pronunciation, whether the languages and accents you need are available, and which audio formats you can download. Pricing is usually per character or per minute, and commercial rights may depend on the plan. AI Voice Over tools add editing features for video narration.
Tools are ordered by one combined score rather than by traffic alone. The largest part is monthly visits to the tool's own website, with GitHub stars used for open-source projects. Smaller parts are the number of Voice Generation & Conversion subcategories the tool covers and how often Creati.ai users save it. Each tool is listed once, and only tools whose main focus is this group qualify. The monthly traffic ranking linked on this page is the traffic-only view.
Building with agents instead? AI Agent categories →