Voice Generation & Conversion 2930

AI-powered audio and video editing software that edits like a document. Descript is an AI-powered audio and video editing software that allows users to edit videos and podcasts like a document. It offers features such as transcription, AI speech, filler word removal, studio sound, eye contact correction, green screen removal, and more. Descript is designed for creators, marketers, and businesses to produce high-quality video and audio content quickly and easily.
AI meeting assistant that records, transcribes, and summarizes meetings across multiple platforms. Fireflies.ai is an AI assistant for meetings that records, transcribes, and allows searching across voice conversations. It uses generative AI to bring ChatGPT to meetings, generating transcripts and smart summaries for platforms like Zoom, Google Meet, and Microsoft Teams. It offers features like comprehensive AI summaries, speaker recognition, conversation intelligence, and integration with various work tools.
All-in-one free AI transcription and summarization tool online Saveto AI is a powerful and all-in-one AI transcription platform designed to convert content into accurate text quickly and easily. It supports videos, audio files, and links from almost any platform, making it highly flexible for different use cases. Beyond transcription, Saveto AI also allows users to generate subtitles, chapters, summaries, translations, and even mind maps from the text. It's built to improve both learning and work efficiency by turning content into structured, usable information.
Audio and video transcription, subtitling, dubbing, and translation services. Happy Scribe provides automatic and human transcription and subtitling services, converting audio and video to text with high accuracy (85-99%) in over 120 languages and 45 formats. It offers AI-powered tools alongside professional language services for transcription, subtitling, dubbing, and translation.
Versatile AI voice generator for text to speech, voiceovers, and translations. Murf AI is a versatile AI voice generator that enables users to convert text to speech with lifelike AI voices. It allows for the creation of studio-quality voiceovers in minutes for podcasts, videos, and professional presentations. With over 200 realistic text-to-speech voices in 20+ languages, Murf simplifies business communication by providing solutions for voiceovers, translations, and various other projects, ensuring clear, engaging, and far-reaching messages.
Peech is a text-to-speech reader converting text to audio in 50+ languages. Peech is a text-to-speech reader that converts text into audio with human-like narration in over 50 languages. It caters to individuals and publishers, offering solutions for converting web articles, e-books, and other texts into audiobooks. Peech supports various input formats and provides AI-powered language detection and voice selection. It aims to make content accessible to a wider audience, including those with dyslexia, ADHD, or vision disabilities.
AI text-to-speech platform with 1500+ lifelike voices, emotion control, and multilingual support. FineVoice is a professional AI voice generator platform specializing in lifelike text-to-speech (TTS) services. It offers over 1,500 realistic AI voices across 154 languages and accents. The platform features advanced TTS models like 'TTS Max,' which supports emotion tags such as happy, sad, and whispering, alongside vocalizations like breathing and laughing. FineVoice also includes tools for voice cloning, real-time voice changing, audio enhancement, and AI-driven content generation, designed to streamline workflows for creators, marketers, and enterprises.
AI Text to Speech, voice cloning, and emotional voice design tool. Noiz AI is an advanced AI Text to Speech (TTS), Voice Clone, and Voice Design Tool designed to create lifelike speech. It allows users to clone voices, precisely control emotional nuance (Emotional TTS), utilize a vast voice library, and perform seamless multilingual dubbing and video translation. Noiz AI also offers developer-ready APIs for integration into various products and platforms.
App for reading text aloud with high-quality voice AI. ElevenReader is an app that reads text aloud using high-quality voice AI. It allows users to listen to free audiobooks and read aloud PDFs, eBooks, and Kindle books.
AI-powered text-to-speech converter with human-like voiceovers and advanced customization options. Voicemaker is an AI-based online Text to Speech converter website that helps content providers, video creators, podcasters, and writers get automated human-like voiceovers. It offers features such as voice effects, pauses, speed, pitch, and volume settings, as well as industry-leading features and a developer API. It has 1.1 million users in over 120 countries and has converted over 100 million characters into voiceovers so far.
Online text-to-speech converter with natural voices and multiple formats support. AnyToSpeech is an online text-to-speech converter that allows users to convert text, PDFs, and URLs into natural-sounding audio. It offers a variety of voices and styles to personalize the audio, and users can listen to the audio immediately. It supports creating audiobooks, MP3s, podcasts, and voiceovers.
Easy online platform for video, image, and GIF editing. Clideo is an online platform that provides a comprehensive suite of tools for easily editing video files, images, and GIFs. It aims to simplify video creation and unlock creativity for users, offering a wide range of functionalities from basic editing to advanced conversions and effects. The platform highlights its accessibility, including free usage options.
AI-powered text-to-speech converter for realistic voiceovers. SpeechGen.io is an AI-powered text-to-speech converter and voice generator that allows users to create realistic voiceovers online. Users can insert any text to generate speech and download audio in MP3 or WAV format for various commercial purposes, including YouTube, TikTok, Instagram, Facebook, Twitch, Twitter, Podcasts, Video Ads, Advertising, E-books, and Presentations. It offers a wide range of natural-sounding voices, custom voice settings, and supports multiple languages.
CapCut is an AI-driven all-in-one video editor and graphic design tool. CapCut is an all-in-one video editor and graphic design tool driven by AI. It offers a range of products including desktop and mobile video editors, and an online creative suite. CapCut provides various video and audio editing tools, text and asset options, and AI magic tools to enhance video creation. It also offers solutions for creativity, lifestyle, and marketing & business needs, along with resources and editing tips.
Free online text-to-speech tool with AI voices and multiple languages. TTSMaker is a free online text-to-speech tool that supports unlimited usage, including commercial use. It offers over 200 AI voices and supports multiple languages. TTSMaker can convert text to speech, allowing users to listen online or download audio files in MP3 or WAV format. It provides various voice styles and settings for customization, such as voice speed, volume, and pitch adjustment.
AI-powered online media tools for video, audio, and photo editing. TopMediai is an AI-powered online platform offering a suite of media tools for video, audio, and photo editing. It caters to content creators with features like text-to-speech, AI cover generation, watermark removal, and more, aiming to revolutionize content production. TopMediai provides simple and efficient AI tools that save time and effort, especially for video creators.
Free online text-to-speech tool with 200+ voices and 70+ languages. Luvvoice is a free online text-to-speech (TTS) tool that turns your text into natural-sounding speech. It offers speech synthesis services and supports multiple languages, with over 200 voices and 70 languages available. Users can convert text to speech online without word limits, listen online, and download files in MP3 format. It also supports file to speech conversion from PDF and TXT formats.
Text-to-speech tool that synthesizes natural speech from short voice samples. Fish Speech is a text-to-speech (TTS) tool developed by the creators of So-VITS-SVC and Bert-VITS2. It can synthesize natural and fluent speech from just 15 seconds of any voice, maintaining the given timbre, style, and accent. Fish Audio is a platform for audio generation, offering various voice models for users to discover and use.
Text-to-speech solution with AI voices for personal, commercial, and educational purposes. NaturalReader is a text-to-speech solution designed for personal, commercial, and educational use. It offers a free online platform, mobile apps, and commercial licenses, utilizing AI voices to read text aloud. It supports multiple languages and provides features like voice cloning and content awareness to enhance the listening experience.
Text-to-speech app for listening to digital content on any device. Speechify is a leading text-to-speech app available on Chrome, iOS, Android, and Mac. It allows users to listen to documents, articles, PDFs, emails, and more. Speechify offers AI voice cloning, AI dubbing, and AI video generation. It is used by millions to hear the internet on any device.
Gladia is a production-ready Speech-to-Text API for teams shipping voice products—high accuracy, multilingual, real-time + async, and add-ons. Gladia is a speech-to-text platform built for production, turning raw audio into structured outputs that power real workflows like meeting summaries, CRM enrichment, contact center QA, and real-time voice assistants. With support for 100+ languages and the ability to handle messy real-world audio—overlapping speakers, accents, code-switching, domain-specific terminology—Gladia is designed for the complexity of actual conversations, not clean studio recordings.
Free online video recording, editing, and AI-powered multimedia service platform. RecCloud is a free multifunctional online application dedicated to providing users with comprehensive video recording and editing services. It offers AI tools including Chatvideo, AI speech-to-text, and AI subtitles. RecCloud is an AI video creation platform, offering free multimedia solutions such as AI video chat, AI subtitles, AI speech-to-text, online screen recording, video editing, storage, and sharing.
Converts audio/video to text, summaries, and insights quickly and accurately. Transcript LOL is a service that converts audio and video into text, summaries, and more in seconds. It offers features like speaker recognition, high accuracy, and the ability to download in multiple formats. It's used for course content, extracting key points from meetings or interviews, and creating social media posts.
AI assistant for audio/video transcription and summaries Tongyi Tingwu is an Alibaba Cloud AI assistant for work and study that helps users transcribe, organize, translate, and summarize audio and video content.