Search by tool name, category, or pricing. 39 tools found.
AI voice generator that creates realistic text-to-speech voiceovers for videos and presentations.
AI voice synthesis platform known for extremely realistic and expressive voice cloning.
Listnr is an AI voice platform for converting text into speech, generating voiceovers, and embedding audio players. Its published product materials also describe voice cloning and API access for speech generation.
WellSaid Labs is an enterprise text-to-speech platform for creating synthetic voiceovers. It offers a studio, API, pronunciation controls, team collaboration, and custom voice avatars for production workflows.
Fliki is a browser-based AI content creation platform that turns text, ideas, presentations, and other source material into videos with synthetic narration, stock media, avatars, captions, and editing tools.
TTSMaker is a browser-based text-to-speech service that converts entered text into synthetic audio. It offers multiple languages and voices, playback and downloadable audio, with free usage subject to service limits.
Resemble AI is a generative voice platform for text-to-speech, speech-to-speech, voice cloning, localization, and deepfake detection. It provides web tools and APIs for creating and managing synthetic voices.
Speechify is a text-to-speech platform that converts documents, webpages, PDFs, and other text into spoken audio. Its products include reader apps, browser extensions, voice generation, dubbing, transcription, and an API.
Typecast is an AI voice and video production platform from South Korea’s Neosapience. It converts scripts into synthetic speech and supports virtual avatars, emotion controls, dubbing, and downloadable media.
NaturalReader is text-to-speech software from Naturalsoft that converts documents, web pages, PDFs, and other text into spoken audio. It offers online, desktop, mobile, commercial, and accessibility-focused products.
Microsoft Azure AI Speech is a cloud service for speech-to-text, text-to-speech, speech translation, speaker recognition, and voice-enabled application development through APIs and SDKs.
Descript Overdub is an AI voice-cloning and text-to-speech feature for creating or correcting spoken audio by editing text. It is integrated into Descript’s audio and video editor and includes consent-based voice controls.
Kits AI is a browser-based music production platform for AI voice conversion, singing voice generation, vocal removal, stem splitting, mastering, and licensed artist voice models.
Altered Studio is a desktop voice-content creation and voice-changing application from Altered Ltd. It supports speech-to-speech voice transformation, text-to-speech, transcription, and audio editing workflows.
Respeecher is a Ukrainian voice-cloning and speech-to-speech platform for producing synthetic dialogue while preserving a performer’s delivery. It serves film, television, game, localization, and accessibility workflows.
Voicemod is real-time voice-changing and soundboard software for Windows and macOS. It provides voice effects, audio routing, custom voice creation, and integrations for games, chat, streaming, and content creation.
iMyFone MagicMic is a real-time voice changer and soundboard for Windows, macOS, iOS, and Android. It provides voice effects, audio-file voice conversion, keybind controls, and integrations aimed at gaming, streaming, and voice chat.
Voice.ai provides real-time AI voice changing and voice generation for gaming, streaming, calls, and content creation. Its ecosystem includes desktop and mobile apps and a library of user-accessible voices.
FineVoice is FineShare's AI voice platform for text-to-speech, voice generation, voice changing, transcription, and audio processing. It is offered through web tools and a Windows desktop application.
Soundful is an AI music-generation platform for creating royalty-free tracks from genre and mood settings. It serves creators, artists, brands, and businesses through web tools, licensing plans, and an API.
Beatoven.ai is an AI music-generation platform for creating customizable background tracks from text prompts. It is aimed at video, podcast, game, and other media creators, with licensing terms for downloaded music.
Loudly is an AI music platform for generating, customizing, and licensing tracks for videos, podcasts, social media, and other digital projects. It also offers music distribution and audio tools for creators.
Stable Audio is Stability AI’s generative-audio product for creating music and sound effects from text prompts. It offers web-based generation and is built around Stability AI’s audio models, including Stable Audio 2.0.
Mubert is a generative music platform that creates royalty-free audio from text prompts and other parameters. It offers tools for creators, developers, listeners, and artists, including an API for integrating generated music.
Riverside is a browser-based platform for recording, editing, and publishing remote podcasts and videos. It captures participants locally, supports separate tracks, transcription, and AI-assisted editing and clip creation.
Alitu is a web-based podcast production tool that automates audio cleanup, leveling, assembly, music insertion, recording, transcription, editing, hosting, and distribution for podcast creators.
Adobe Podcast is a browser-based suite for recording and editing spoken audio. Its AI features include speech enhancement, microphone checks, transcription-based editing, and remote recording tools.
Castmagic is an AI content workspace that transcribes audio and video, then generates summaries, show notes, timestamps, quotes, social posts, newsletters, and other reusable assets from recordings.
Auphonic is a web-based and desktop audio post-production service that applies loudness normalization, leveling, noise and hum reduction, filtering, encoding, metadata, and publishing workflows.
Cleanvoice AI is an online audio-editing service that automatically removes filler words, mouth sounds, stutters, long silences, and background noise from podcast and spoken-word recordings.
Podcastle is a browser-based audio and video creation platform for recording, editing, transcription, text-to-speech, voice cloning, and podcast production, with tools aimed at creators and teams.
Swell AI is a content repurposing platform for podcasts and videos. It uses AI to generate transcripts, show notes, articles, clips, social posts, newsletters, titles, and other derivative media.
Podsqueeze is an AI podcast post-production platform that turns uploaded audio or video into transcripts, show notes, timestamps, clips, newsletters, social posts, titles, and related publishing assets.
Ecrett Music is a browser-based music generator for videos, podcasts, games, and other media. Users select a scene, mood, and genre, then customize generated tracks and license downloads under a paid plan.
Speechelo is a cloud-based text-to-speech application from Blaster Suite. It converts entered text into synthetic voice-over audio, with multiple voices, languages, and speech-style controls for video narration.
Simon Says is an AI-assisted transcription and subtitle platform for audio and video. It provides automated transcription, translation, captioning, browser-based editing, and integrations for professional video-editing workflows.
Replica Studios was an AI voice platform for generating character dialogue and speech for games, animation, and other media. Its website now states that the service has shut down and directs visitors to support resources.
Udio launched in April 2024 from ex-DeepMind researchers and raised $10M from a16z, will.i.am, and Common. It generates studio-quality vocal tracks from text prompts and quickly became Suno's closest competitor. The RIAA sued Udio in June 2024 over alleged training on copyrighted recordings. Despite the litigation, Udio has continued shipping new models, including v1.5 with full-song generation and remixing.
Suno launched in late 2023 as a text-to-song generator that creates full vocal tracks with melody, lyrics, and instrumentation from short prompts. It raised $125M at a $500M valuation in May 2024 and crossed 12M users by mid-2024. In June 2024 the RIAA sued Suno (and competitor Udio) for copyright infringement on behalf of Sony, Universal, and Warner. The product keeps shipping new versions and signups remain open while the litigation runs.