AI Text-to-Speech 386

AI storybook generator with personalized characters and illustrations.
Childbook.ai is an AI Story Book Generator that allows users to create stunning AI-generated children's books with personalized characters and unique illustrations. It caters to parents, teachers, and storytellers, enabling them to transform their stories into beautiful books. Users can add their photo to become the main character, create stories in any language, edit illustrations, rewrite plots, and even listen to their books with synchronized text or order printed copies.

Browser plugin for translating and dubbing foreign videos on YouTube and other platforms.
YouTube Dubbing is a browser plugin that translates foreign subtitles and reads them aloud, enhancing the viewing experience by eliminating language barriers. It combines AI subtitle translation and online dubbing, supporting GPT and Claude models. Compatible with multiple browsers and devices, it works on platforms such as YouTube, Udemy, Bilibili, and gamedev.tv. Key features include multiple voice options, playback speed control, background sound preservation, speaker recognition, translated subtitle display, and webpage text dubbing. It offers both free and membership modes for enhanced features.

Perso Dubbing is an AI video dubbing platform that translates, dubs, and lip-syncs videos into 99+ languages. AI voice cloning preserves each speaker's tone and emotion, and multi-speaker detection handles up to 10 speakers per video. It reduces localization costs by up to 98% compared to traditional dubbing studios. Developed by ESTsoft and trusted by 450,000+ users.
Perso Dubbing is an AI-powered video dubbing and translation platform that localizes content into 99+ languages in minutes, with speech recognition in 100+ languages. Teams upload a video, select target languages, and receive a studio-quality dubbed version — complete with lip-sync and voice cloning that preserves the original speaker's tone, accent, and emotion.
Key capabilities:
• AI Voice Cloning — Matches the original speaker's voice, accent, and emotional tone across all dubbed tracks
• AI Lip Sync — Aligns translated audio with on-screen mouth movements for natural viewing
• Speech-to-Text — Speech recognition in 100+ languages
• Audio Separation — Splits voice and background tracks
• Auto Subtitle Generation — Creates and exports subtitles automatically
• Real-Time Script Editor — Review and refine translations before final export
• Multi-Speaker Support — Detects and dubs up to 10 speakers in a single workflow
Built for marketing teams, e-learning creators, enterprise L&D departments, and media publishers expanding into global markets. Enterprise plans include API access, advanced security controls, and dedicated support. Developed by ESTsoft (est. 1993, KOSDAQ: 047560) — ISO/IEC 27001 and KISA ISMS certified.

AI-powered video localization and dubbing tool for global content expansion.
Rask AI is an AI-powered video localization and dubbing tool designed to provide human-quality dubbing and translation experiences. It offers features such as video translation, transcription, lip-syncing, and voice cloning in multiple languages. The platform aims to help businesses and creators expand their global reach by automatically translating and dubbing their content, including marketing videos, podcasts, and lectures, into over 130 languages.

AI-powered multilingual voice synthesis and cloning platform with natural language processing.
VoiceCanvas is an advanced AI-powered multilingual voice synthesis and voice cloning platform. It offers state-of-the-art neural voice synthesis and voice cloning technology in over 40 languages, providing clear and transparent audio quality, natural language processing, and personalized voice cloning features. It is a professional-grade text-to-speech platform with advanced AI technology.

AI platform converts novels to audiobooks with unique character voices.
VoiceNovel is an advanced AI voice synthesis platform that transforms novels into high-quality voice novels and audiobooks. It leverages AI technology to convert text into natural-sounding speech, supporting multiple voice styles to give each character a unique voice and create an immersive listening experience. The platform offers features for novel upload and analysis, a personal library for converted audiobooks, and an audio player with download options for premium users.

Free browser tool that reads text aloud with natural AI voices
Read Aloud Reader is a free browser-based text-to-speech tool that converts typed, pasted, or uploaded content into natural-sounding speech. Users can listen to articles, PDFs, DOCX files, EPUBs, notes, emails, and other text with neural AI voices, sentence highlighting, adjustable playback speed, and MP3 export. It works across desktop and mobile browsers without requiring installation or an account.

Text to speech platform using OpenAI technology for high-quality audio conversion.
Turn your PDFs and eBooks into exciting AudioBooks or MP3 files. Perfect for creating quick spoken books, podcasts for learning anytime, anywhere: while driving, exercising, or relaxing.. Discover the future of digital communication with our cutting-edge Text To Speech OpenAI technology. Our advanced Voice Engine transforms text into natural-sounding speech, seamlessly bridging the gap between humans and machines. Ideal for developers, creators, and businesses, our platform offers an intuitive API for easy integration, ensuring your applications and services are more accessible and engaging than ever. Experience unparalleled voice quality and flexibility with our solution, designed to elevate your digital interactions to new heights

Open-weights 8B AI text-to-speech model for expressive English speech.
Miso One is an open-weights, 8B-parameter text-to-speech (TTS) system developed by Miso Labs. It is designed specifically for producing highly realistic, expressive, and emotionally varied English conversational speech, making it ideal for voice-agent research and developer workflows. Built on a Sesame-style conversational speech model (CSM) architecture with Mimi audio codes, it features a highly optimized inference capability boasting a published low latency of 110 ms. In addition to text-to-speech generation, the model supports voice continuation and one-shot voice cloning from audio context with clear consent boundaries.

DesiVocal is a free AI voice generator for HD voice overs in multiple languages.
DesiVocal is a free text-to-speech and AI voice generator that creates HD AI voice overs in multiple languages. It caters to youtubers, publishers, and media houses, offering premium AI voice overs in seconds. It also provides a speech-to-text feature.

AI-powered text-to-speech system with natural speech, voice cloning, and multi-language support.
F5-TTS is an advanced AI-powered text-to-speech system that converts text into natural, expressive speech. It supports multi-language synthesis, emotional control, and speed adjustments, making it perfect for audiobooks, assistants, and content creation. F5-TTS offers zero-shot voice cloning, multi-language support, and emotion expression capabilities.

Text-to-speech reader for webpages, PDFs, Kindle books, and AI answers
CastReader is a text-to-speech and reading-assistance tool available as a browser extension and mobile app. It reads webpages, PDFs, Kindle books, ebooks, documents, emails, and AI responses aloud with synchronized highlighting and auto-scroll. Its Read & Explain feature helps users understand difficult passages through spoken explanations, subtitles, and annotations while keeping the original source visible. CastReader supports multiple languages, including German, and works across supported Chrome, Edge, Firefox, iPhone, iPad, and Android workflows.

AI text-to-speech platform with 800+ voices for content creation and more.
SteosVoice (formerly CyberVoice) is an AI-powered text-to-speech platform that offers over 800 voices for speech synthesis. It allows users to convert text into high-quality audio for various applications, including YouTube localization, content creation, mods, audiobooks, and more. The platform provides both free and paid options, including a Telegram bot for free limited access and subscription plans for more extensive use.

Text-to-speech Chrome extension for reading aloud digital content in multiple languages.
Voice Out is a text-to-speech Chrome extension that reads aloud Google Docs, PDFs, webpages, or books in 60+ languages with 100+ voices. It's designed to be fast, easy, and free, allowing users to listen to content while browsing, working, or relaxing.

Classic Microsoft SAM Text-to-Speech voice in your browser.
Microsoft SAM Text-to-Speech is a modern JavaScript implementation of the iconic voice synthesizer from Windows XP, originally part of the Microsoft Speech API (SAPI). This website brings the classic Microsoft SAM voice directly to your browser, allowing users to generate speech with its distinctive robotic voice without any downloads or server processing. It aims to preserve the authentic nostalgic charm of the original while adding modern conveniences like browser-based functionality and customizable parameters.

AI voice generator with 300 voices in 70+ languages for lifelike speech synthesis.
Lovevoice AI Voice Generator transforms text into lifelike speech using AI technology. It offers nearly 300 AI voices in over 70 languages, allowing users to create natural-sounding audio for various applications, including videos, podcasts, audiobooks, presentations, and marketing materials. Users can adjust speed, volume, and pitch to customize the generated voices. The service supports multiple file formats for transcription and processes large volumes of text quickly.

Free online AI text to speech generator with realistic voices and customization.
PopPop AI Text to Speech is a free online AI voice and speech generation tool that offers over 200 characters in 20+ languages. It provides fast, natural speech generated by AI without ads or signup requirements. The tool allows users to convert text to audio using realistic AI voices and customize the speed and pitch of the voice.

Free AI text-to-speech for 140+ languages and MP3 downloads
TTSFree is a free online text-to-speech website that converts written text into natural-sounding AI voices in 140+ languages, with MP3 downloads and customization options.

AI-powered platform for creating logos, videos, and designs quickly and easily.
Designs AI is an all-in-one AI-powered creative platform that allows users to create logos, videos, banners, mockups, and more in minutes. It offers tools like an AI design generator, image generator, video maker, logo generator, AI Writer, and AI Chat to streamline content creation and marketing efforts. The platform aims to make design accessible to everyone, regardless of their experience level.

AI-powered platform for automated podcast and audio content generation from multiple sources.
AutoContent API is an AI-powered platform that generates podcasts and other audio content from various sources like websites, text, and YouTube videos. It offers a comprehensive solution for automated podcast generation, including multilanguage support, multi-voice generation, and custom voice options, designed for professional content creators.

PodcastorAI is an all-in-one podcast production platform — from source material to published episode, without a studio, camera, or editing software.
PodcastorAI is an all-in-one podcast production platform that takes you from source material to a published audio or video episode without a studio, a camera, or a co-producer. Upload a topic, URL, document, or audio file to generate a structured co-host script, produce AI audio using voices powered by ElevenLabs and MiniMax, and add an AI avatar as your on-screen host across multiple video formats. Script, audio, video, subtitles, and publishing all stay inside the same workflow.

AI-powered platform for studio-quality video and podcast creation, editing, and distribution.
Podcastle is the easiest way to create studio-quality videos and podcasts. Record, edit and distribute content directly in your browser using AI-powered tools. It's a one-stop shop for broadcast storytelling, great for podcasters or anyone who deals with long-form video creation. Studio-quality recording, AI-powered editing, and seamless exporting – all in a single web-based platform.

AI-powered podcast builder for effortless content repurposing and creation.
Wondercraft is a podcast builder that leverages AI voices to let anyone go from idea to published podcast in minutes. It allows users to repurpose existing content (newsletters, blogs, interviews, recordings) to create engaging podcasts effortlessly.

AI audio tour guide with AR, providing local stories and events.
Summer AI is an AI audio tour guide that provides nearby stories, points of interest, and local events. It uses LLMs to summarize information and TTS to read it out. It also features an Augmented Reality mode.
Hot Articles
How to Build a Multimodal AI Knowledge Base With Gemini Embedding 2
NASA 和 IBM 开源月球模型
Introducing the Australian Youth Safety Blueprint
GLM-5.3-FlashX 上线,智谱把国产卡上的推理速度顶到 200 token/s
Google’s Gemini is the latest AI model to hack other companies
How to Choose Hardware for Running Local LLMs, and Know Exactly When It Beats the Claude API
Latest Articles
Google’s Gemini is the latest AI model to hack other companies
How to Choose Hardware for Running Local LLMs, and Know Exactly When It Beats the Claude API
How to Build a Multimodal AI Knowledge Base With Gemini Embedding 2
Announcing Grok-1.5
Google DeepMind launches institute to widen the AGI debate
Google’s new ‘CC’ is an AI agent that helps families run their households
Hot Tags