Search by tool name, category, or pricing. 50 tools found.
AI text-to-speech platform with ultra-realistic voices for podcasts, videos, and audiobooks.
AI voice generator that creates realistic text-to-speech voiceovers for videos and presentations.
AI voice synthesis platform known for extremely realistic and expressive voice cloning.
Listnr is an AI voice platform for converting text into speech, generating voiceovers, and embedding audio players. Its published product materials also describe voice cloning and API access for speech generation.
WellSaid Labs is an enterprise text-to-speech platform for creating synthetic voiceovers. It offers a studio, API, pronunciation controls, team collaboration, and custom voice avatars for production workflows.
Fliki is a browser-based AI content creation platform that turns text, ideas, presentations, and other source material into videos with synthetic narration, stock media, avatars, captions, and editing tools.
TTSMaker is a browser-based text-to-speech service that converts entered text into synthetic audio. It offers multiple languages and voices, playback and downloadable audio, with free usage subject to service limits.
Resemble AI is a generative voice platform for text-to-speech, speech-to-speech, voice cloning, localization, and deepfake detection. It provides web tools and APIs for creating and managing synthetic voices.
Speechify is a text-to-speech platform that converts documents, webpages, PDFs, and other text into spoken audio. Its products include reader apps, browser extensions, voice generation, dubbing, transcription, and an API.
Typecast is an AI voice and video production platform from South Korea’s Neosapience. It converts scripts into synthetic speech and supports virtual avatars, emotion controls, dubbing, and downloadable media.
NaturalReader is text-to-speech software from Naturalsoft that converts documents, web pages, PDFs, and other text into spoken audio. It offers online, desktop, mobile, commercial, and accessibility-focused products.
Microsoft Azure AI Speech is a cloud service for speech-to-text, text-to-speech, speech translation, speaker recognition, and voice-enabled application development through APIs and SDKs.
Lyrebird developed AI systems for copying a person’s voice from short audio samples and generating speech in that voice. Descript acquired the Montreal startup in 2019 and incorporated its technology and team.
Descript Overdub is an AI voice-cloning and text-to-speech feature for creating or correcting spoken audio by editing text. It is integrated into Descript’s audio and video editor and includes consent-based voice controls.
Kits AI is a browser-based music production platform for AI voice conversion, singing voice generation, vocal removal, stem splitting, mastering, and licensed artist voice models.
Altered Studio is a desktop voice-content creation and voice-changing application from Altered Ltd. It supports speech-to-speech voice transformation, text-to-speech, transcription, and audio editing workflows.
Respeecher is a Ukrainian voice-cloning and speech-to-speech platform for producing synthetic dialogue while preserving a performer’s delivery. It serves film, television, game, localization, and accessibility workflows.
Voicemod is real-time voice-changing and soundboard software for Windows and macOS. It provides voice effects, audio routing, custom voice creation, and integrations for games, chat, streaming, and content creation.
iMyFone MagicMic is a real-time voice changer and soundboard for Windows, macOS, iOS, and Android. It provides voice effects, audio-file voice conversion, keybind controls, and integrations aimed at gaming, streaming, and voice chat.
Voice.ai provides real-time AI voice changing and voice generation for gaming, streaming, calls, and content creation. Its ecosystem includes desktop and mobile apps and a library of user-accessible voices.
VocaliD developed personalized synthetic voices for people who could not speak and later offered custom voice technology for brands. The company combined donated speech recordings with a recipient’s vocal characteristics.
FineVoice is FineShare's AI voice platform for text-to-speech, voice generation, voice changing, transcription, and audio processing. It is offered through web tools and a Windows desktop application.
Amper Music was an AI-assisted music composition platform that let users generate and customize royalty-cleared tracks for videos, podcasts, games, and other media. Shutterstock acquired the company in 2020.
Soundful is an AI music-generation platform for creating royalty-free tracks from genre and mood settings. It serves creators, artists, brands, and businesses through web tools, licensing plans, and an API.
Beatoven.ai is an AI music-generation platform for creating customizable background tracks from text prompts. It is aimed at video, podcast, game, and other media creators, with licensing terms for downloaded music.
Loudly is an AI music platform for generating, customizing, and licensing tracks for videos, podcasts, social media, and other digital projects. It also offers music distribution and audio tools for creators.
Stable Audio is Stability AI’s generative-audio product for creating music and sound effects from text prompts. It offers web-based generation and is built around Stability AI’s audio models, including Stable Audio 2.0.
Mubert is a generative music platform that creates royalty-free audio from text prompts and other parameters. It offers tools for creators, developers, listeners, and artists, including an API for integrating generated music.
Riverside is a browser-based platform for recording, editing, and publishing remote podcasts and videos. It captures participants locally, supports separate tracks, transcription, and AI-assisted editing and clip creation.
Alitu is a web-based podcast production tool that automates audio cleanup, leveling, assembly, music insertion, recording, transcription, editing, hosting, and distribution for podcast creators.
Adobe Podcast is a browser-based suite for recording and editing spoken audio. Its AI features include speech enhancement, microphone checks, transcription-based editing, and remote recording tools.
Castmagic is an AI content workspace that transcribes audio and video, then generates summaries, show notes, timestamps, quotes, social posts, newsletters, and other reusable assets from recordings.
Auphonic is a web-based and desktop audio post-production service that applies loudness normalization, leveling, noise and hum reduction, filtering, encoding, metadata, and publishing workflows.
Cleanvoice AI is an online audio-editing service that automatically removes filler words, mouth sounds, stutters, long silences, and background noise from podcast and spoken-word recordings.
Podcastle is a browser-based audio and video creation platform for recording, editing, transcription, text-to-speech, voice cloning, and podcast production, with tools aimed at creators and teams.
Swell AI is a content repurposing platform for podcasts and videos. It uses AI to generate transcripts, show notes, articles, clips, social posts, newsletters, titles, and other derivative media.
Podsqueeze is an AI podcast post-production platform that turns uploaded audio or video into transcripts, show notes, timestamps, clips, newsletters, social posts, titles, and related publishing assets.
Ecrett Music is a browser-based music generator for videos, podcasts, games, and other media. Users select a scene, mood, and genre, then customize generated tracks and license downloads under a paid plan.
Speechelo is a cloud-based text-to-speech application from Blaster Suite. It converts entered text into synthetic voice-over audio, with multiple voices, languages, and speech-style controls for video narration.
Simon Says is an AI-assisted transcription and subtitle platform for audio and video. It provides automated transcription, translation, captioning, browser-based editing, and integrations for professional video-editing workflows.
Replica Studios was an AI voice platform for generating character dialogue and speech for games, animation, and other media. Its website now states that the service has shut down and directs visitors to support resources.
Jukedeck was a London-based AI music platform that generated royalty-free tracks from user-selected parameters such as genre, mood, duration, and tempo. TikTok parent ByteDance acquired the company in 2019.
LOVO AI is a generative voice and speech platform for creating text-to-speech audio, voiceovers, voice clones, and narrated videos. Its Genny application combines AI voice generation with script, subtitle, image, and video editing tools.
Veritone Voice is a synthetic-voice platform for creating, managing, and licensing AI-generated voice content. It supports custom voice models and is aimed at media, entertainment, sports, advertising, and related uses.
Humtap was an AI-assisted mobile music creation app that turned hummed melodies and tapped rhythms into arranged songs. Its developer later shifted focus to enterprise generative media technology under the MWM brand.
Altered was a real-time AI voice-changing product from Modulate that let users select synthetic character voices for online games and voice chat. Modulate later shifted its public focus to ToxMod, its voice-safety platform.
Melodrive was an adaptive music system that generated and adjusted music in real time for interactive media, including games and virtual-reality experiences, based on emotional parameters.
Podium was an AI post-production platform for podcasters that generated transcripts, show notes, chapters, clips, highlights, and social posts from uploaded audio. The service’s present operating status is unclear.
Udio launched in April 2024 from ex-DeepMind researchers and raised $10M from a16z, will.i.am, and Common. It generates studio-quality vocal tracks from text prompts and quickly became Suno's closest competitor. The RIAA sued Udio in June 2024 over alleged training on copyrighted recordings. Despite the litigation, Udio has continued shipping new models, including v1.5 with full-song generation and remixing.
Suno launched in late 2023 as a text-to-song generator that creates full vocal tracks with melody, lyrics, and instrumentation from short prompts. It raised $125M at a $500M valuation in May 2024 and crossed 12M users by mid-2024. In June 2024 the RIAA sued Suno (and competitor Udio) for copyright infringement on behalf of Sony, Universal, and Warner. The product keeps shipping new versions and signups remain open while the litigation runs.