23 tools found
Castmagic is an AI content workspace that transcribes audio and video, then generates summaries, show notes, timestamps, quotes, social posts, newsletters, and other reusable assets from recordings.
Swell AI is a content repurposing platform for podcasts and videos. It uses AI to generate transcripts, show notes, articles, clips, social posts, newsletters, titles, and other derivative media.
Podcastle is a browser-based audio and video creation platform for recording, editing, transcription, text-to-speech, voice cloning, and podcast production, with tools aimed at creators and teams.
Mubert is a generative music platform that creates royalty-free audio from text prompts and other parameters. It offers tools for creators, developers, listeners, and artists, including an API for integrating generated music.
Stable Audio is Stability AI’s generative-audio product for creating music and sound effects from text prompts. It offers web-based generation and is built around Stability AI’s audio models, including Stable Audio 2.0.
Loudly is an AI music platform for generating, customizing, and licensing tracks for videos, podcasts, social media, and other digital projects. It also offers music distribution and audio tools for creators.
Beatoven.ai is an AI music-generation platform for creating customizable background tracks from text prompts. It is aimed at video, podcast, game, and other media creators, with licensing terms for downloaded music.
FineVoice is FineShare's AI voice platform for text-to-speech, voice generation, voice changing, transcription, and audio processing. It is offered through web tools and a Windows desktop application.
Respeecher is a Ukrainian voice-cloning and speech-to-speech platform for producing synthetic dialogue while preserving a performer’s delivery. It serves film, television, game, localization, and accessibility workflows.
Voice.ai provides real-time AI voice changing and voice generation for gaming, streaming, calls, and content creation. Its ecosystem includes desktop and mobile apps and a library of user-accessible voices.
Altered Studio is a desktop voice-content creation and voice-changing application from Altered Ltd. It supports speech-to-speech voice transformation, text-to-speech, transcription, and audio editing workflows.
Descript Overdub is an AI voice-cloning and text-to-speech feature for creating or correcting spoken audio by editing text. It is integrated into Descript’s audio and video editor and includes consent-based voice controls.
Microsoft Azure AI Speech is a cloud service for speech-to-text, text-to-speech, speech translation, speaker recognition, and voice-enabled application development through APIs and SDKs.
NaturalReader is text-to-speech software from Naturalsoft that converts documents, web pages, PDFs, and other text into spoken audio. It offers online, desktop, mobile, commercial, and accessibility-focused products.
Typecast is an AI voice and video production platform from South Korea’s Neosapience. It converts scripts into synthetic speech and supports virtual avatars, emotion controls, dubbing, and downloadable media.
Speechify is a text-to-speech platform that converts documents, webpages, PDFs, and other text into spoken audio. Its products include reader apps, browser extensions, voice generation, dubbing, transcription, and an API.
Resemble AI is a generative voice platform for text-to-speech, speech-to-speech, voice cloning, localization, and deepfake detection. It provides web tools and APIs for creating and managing synthetic voices.
TTSMaker is a browser-based text-to-speech service that converts entered text into synthetic audio. It offers multiple languages and voices, playback and downloadable audio, with free usage subject to service limits.
Fliki is a browser-based AI content creation platform that turns text, ideas, presentations, and other source material into videos with synthetic narration, stock media, avatars, captions, and editing tools.
WellSaid Labs is an enterprise text-to-speech platform for creating synthetic voiceovers. It offers a studio, API, pronunciation controls, team collaboration, and custom voice avatars for production workflows.
Listnr is an AI voice platform for converting text into speech, generating voiceovers, and embedding audio players. Its published product materials also describe voice cloning and API access for speech generation.
AI voice generator that creates realistic text-to-speech voiceovers for videos and presentations.
AI voice synthesis platform known for extremely realistic and expressive voice cloning.