Voice Generation & Conversion 3019

Perso Dubbing is an AI video dubbing platform that translates, dubs, and lip-syncs videos into 99+ languages. AI voice cloning preserves each speaker's tone and emotion, and multi-speaker detection handles up to 10 speakers per video. It reduces localization costs by up to 98% compared to traditional dubbing studios. Developed by ESTsoft and trusted by 450,000+ users. Perso Dubbing is an AI-powered video dubbing and translation platform that localizes content into 99+ languages in minutes, with speech recognition in 100+ languages. Teams upload a video, select target languages, and receive a studio-quality dubbed version — complete with lip-sync and voice cloning that preserves the original speaker's tone, accent, and emotion. Key capabilities: • AI Voice Cloning — Matches the original speaker's voice, accent, and emotional tone across all dubbed tracks • AI Lip Sync — Aligns translated audio with on-screen mouth movements for natural viewing • Speech-to-Text — Speech recognition in 100+ languages • Audio Separation — Splits voice and background tracks • Auto Subtitle Generation — Creates and exports subtitles automatically • Real-Time Script Editor — Review and refine translations before final export • Multi-Speaker Support — Detects and dubs up to 10 speakers in a single workflow Built for marketing teams, e-learning creators, enterprise L&D departments, and media publishers expanding into global markets. Enterprise plans include API access, advanced security controls, and dedicated support. Developed by ESTsoft (est. 1993, KOSDAQ: 047560) — ISO/IEC 27001 and KISA ISMS certified.
AI-powered video localization and dubbing tool for global content expansion. Rask AI is an AI-powered video localization and dubbing tool designed to provide human-quality dubbing and translation experiences. It offers features such as video translation, transcription, lip-syncing, and voice cloning in multiple languages. The platform aims to help businesses and creators expand their global reach by automatically translating and dubbing their content, including marketing videos, podcasts, and lectures, into over 130 languages.
AI platform to generate, edit, and translate talking videos with prompts. Vozo is an AI-powered platform that enables users to generate, edit, and translate talking videos with prompts. It allows for rewriting, redubbing, voice editing, and lip-syncing of existing videos. Users can transform classics into promos, ordinary videos into comedies, or translate content into multiple languages. It also offers features like auto subtitles, voice changing, and voiceover modification.
AI-powered multilingual voice synthesis and cloning platform with natural language processing. VoiceCanvas is an advanced AI-powered multilingual voice synthesis and voice cloning platform. It offers state-of-the-art neural voice synthesis and voice cloning technology in over 40 languages, providing clear and transparent audio quality, natural language processing, and personalized voice cloning features. It is a professional-grade text-to-speech platform with advanced AI technology.
AI platform converts novels to audiobooks with unique character voices. VoiceNovel is an advanced AI voice synthesis platform that transforms novels into high-quality voice novels and audiobooks. It leverages AI technology to convert text into natural-sounding speech, supporting multiple voice styles to give each character a unique voice and create an immersive listening experience. The platform offers features for novel upload and analysis, a personal library for converted audiobooks, and an audio player with download options for premium users.
Free browser tool that reads text aloud with natural AI voices Read Aloud Reader is a free browser-based text-to-speech tool that converts typed, pasted, or uploaded content into natural-sounding speech. Users can listen to articles, PDFs, DOCX files, EPUBs, notes, emails, and other text with neural AI voices, sentence highlighting, adjustable playback speed, and MP3 export. It works across desktop and mobile browsers without requiring installation or an account.
Text to speech platform using OpenAI technology for high-quality audio conversion. Turn your PDFs and eBooks into exciting AudioBooks or MP3 files. Perfect for creating quick spoken books, podcasts for learning anytime, anywhere: while driving, exercising, or relaxing.. Discover the future of digital communication with our cutting-edge Text To Speech OpenAI technology. Our advanced Voice Engine transforms text into natural-sounding speech, seamlessly bridging the gap between humans and machines. Ideal for developers, creators, and businesses, our platform offers an intuitive API for easy integration, ensuring your applications and services are more accessible and engaging than ever. Experience unparalleled voice quality and flexibility with our solution, designed to elevate your digital interactions to new heights
Open-weights 8B AI text-to-speech model for expressive English speech. Miso One is an open-weights, 8B-parameter text-to-speech (TTS) system developed by Miso Labs. It is designed specifically for producing highly realistic, expressive, and emotionally varied English conversational speech, making it ideal for voice-agent research and developer workflows. Built on a Sesame-style conversational speech model (CSM) architecture with Mimi audio codes, it features a highly optimized inference capability boasting a published low latency of 110 ms. In addition to text-to-speech generation, the model supports voice continuation and one-shot voice cloning from audio context with clear consent boundaries.
DesiVocal is a free AI voice generator for HD voice overs in multiple languages. DesiVocal is a free text-to-speech and AI voice generator that creates HD AI voice overs in multiple languages. It caters to youtubers, publishers, and media houses, offering premium AI voice overs in seconds. It also provides a speech-to-text feature.
AI-powered text-to-speech system with natural speech, voice cloning, and multi-language support. F5-TTS is an advanced AI-powered text-to-speech system that converts text into natural, expressive speech. It supports multi-language synthesis, emotional control, and speed adjustments, making it perfect for audiobooks, assistants, and content creation. F5-TTS offers zero-shot voice cloning, multi-language support, and emotion expression capabilities.
Text-to-speech reader for webpages, PDFs, Kindle books, and AI answers CastReader is a text-to-speech and reading-assistance tool available as a browser extension and mobile app. It reads webpages, PDFs, Kindle books, ebooks, documents, emails, and AI responses aloud with synchronized highlighting and auto-scroll. Its Read & Explain feature helps users understand difficult passages through spoken explanations, subtitles, and annotations while keeping the original source visible. CastReader supports multiple languages, including German, and works across supported Chrome, Edge, Firefox, iPhone, iPad, and Android workflows.
AI text-to-speech platform with 800+ voices for content creation and more. SteosVoice (formerly CyberVoice) is an AI-powered text-to-speech platform that offers over 800 voices for speech synthesis. It allows users to convert text into high-quality audio for various applications, including YouTube localization, content creation, mods, audiobooks, and more. The platform provides both free and paid options, including a Telegram bot for free limited access and subscription plans for more extensive use.
Text-to-speech Chrome extension for reading aloud digital content in multiple languages. Voice Out is a text-to-speech Chrome extension that reads aloud Google Docs, PDFs, webpages, or books in 60+ languages with 100+ voices. It's designed to be fast, easy, and free, allowing users to listen to content while browsing, working, or relaxing.
AI voice generator with realistic text-to-speech and speech-to-speech capabilities. Respeecher Voice Marketplace is an AI voice generator platform that offers realistic text-to-speech and speech-to-speech capabilities. It provides a range of AI voice solutions for creative and professional projects, including film and TV production, game development, advertising, and more. The platform is trusted by industry leaders and offers high-quality AI voices, including celebrity voices, with a focus on ethical use and legal compliance.
Classic Microsoft SAM Text-to-Speech voice in your browser. Microsoft SAM Text-to-Speech is a modern JavaScript implementation of the iconic voice synthesizer from Windows XP, originally part of the Microsoft Speech API (SAPI). This website brings the classic Microsoft SAM voice directly to your browser, allowing users to generate speech with its distinctive robotic voice without any downloads or server processing. It aims to preserve the authentic nostalgic charm of the original while adding modern conveniences like browser-based functionality and customizable parameters.
AI voice generator with 300 voices in 70+ languages for lifelike speech synthesis. Lovevoice AI Voice Generator transforms text into lifelike speech using AI technology. It offers nearly 300 AI voices in over 70 languages, allowing users to create natural-sounding audio for various applications, including videos, podcasts, audiobooks, presentations, and marketing materials. Users can adjust speed, volume, and pitch to customize the generated voices. The service supports multiple file formats for transcription and processes large volumes of text quickly.
Free online AI text to speech generator with realistic voices and customization. PopPop AI Text to Speech is a free online AI voice and speech generation tool that offers over 200 characters in 20+ languages. It provides fast, natural speech generated by AI without ads or signup requirements. The tool allows users to convert text to audio using realistic AI voices and customize the speed and pitch of the voice.
Free AI text-to-speech for 140+ languages and MP3 downloads TTSFree is a free online text-to-speech website that converts written text into natural-sounding AI voices in 140+ languages, with MP3 downloads and customization options.
AI-powered platform for creating logos, videos, and designs quickly and easily. Designs AI is an all-in-one AI-powered creative platform that allows users to create logos, videos, banners, mockups, and more in minutes. It offers tools like an AI design generator, image generator, video maker, logo generator, AI Writer, and AI Chat to streamline content creation and marketing efforts. The platform aims to make design accessible to everyone, regardless of their experience level.
Modern podcast hosting and distribution platform with unlimited features and AI tools. Podhome is a modern podcast hosting and distribution platform offering unlimited shows, episodes, uploads, and downloads for a monthly price. It includes features like dynamic audio & text, a website per show, and Podhome AI to automatically generate transcripts, chapters, clips, show-notes, and titles. Podhome also provides analytics, easy distribution to podcast directories, audio enhancement, team collaboration tools, embeddable players, and support for Podcasting 2.0 features.
Personalized audio intelligence transforming daily information into interactive audio content. Huxe transforms daily information into personalized audio intelligence, allowing users to stay informed without endless scrolling. It offers a 24/7 live station based on individual interests like neighborhood news, stock portfolios, or favorite sports teams. Huxe also provides a personalized audio briefing for mornings, replacing the need to bounce between multiple apps. Users can interact with the audio, asking for different explanations or more technical details. Additionally, it can turn any curiosity into a personal podcast, providing clear audio explanations.
AI-powered platform for automated podcast and audio content generation from multiple sources. AutoContent API is an AI-powered platform that generates podcasts and other audio content from various sources like websites, text, and YouTube videos. It offers a comprehensive solution for automated podcast generation, including multilanguage support, multi-voice generation, and custom voice options, designed for professional content creators.
PodcastorAI is an all-in-one podcast production platform — from source material to published episode, without a studio, camera, or editing software. PodcastorAI is an all-in-one podcast production platform that takes you from source material to a published audio or video episode without a studio, a camera, or a co-producer. Upload a topic, URL, document, or audio file to generate a structured co-host script, produce AI audio using voices powered by ElevenLabs and MiniMax, and add an AI avatar as your on-screen host across multiple video formats. Script, audio, video, subtitles, and publishing all stay inside the same workflow.
AI-powered platform for studio-quality video and podcast creation, editing, and distribution. Podcastle is the easiest way to create studio-quality videos and podcasts. Record, edit and distribute content directly in your browser using AI-powered tools. It's a one-stop shop for broadcast storytelling, great for podcasters or anyone who deals with long-form video creation. Studio-quality recording, AI-powered editing, and seamless exporting – all in a single web-based platform.