AI Voice Cloning 212

AudioBook Bot uses AI to convert text to audiobooks with multiple voices.
AudioBook Bot converts written works to audio works using text to speech AI. Using AudioBook Bot, you can create character rich audiobooks using multiple licensed voices or read books and podcasts in your own voice. It is a one-click Audiobook creation software that uses generative AI to convert text to speech. With some additional annotations, it can provide your book with a whole cast of characters. You can also narrate the book in your own voice with a minimal sample.
AI platform to convert text files into human-like voiceovers with voice cloning.
Cugent is an AI-powered platform that turns PDF, Doc, and text files into MP3 voiceovers. It supports different voice types in all supported languages and offers multilingual voice types for German, Dutch, Italian, and French scripts. It also features voice cloning, allowing users to use their own voice in all supported languages.

AI dubbing and voiceover tool for media and entertainment with cost-effective localization.
Dubformer is an AI dubbing and voiceover tool for media and entertainment, offering monetization-ready dubbing that sounds like real humans. It provides cost-effective AI localization solutions, including best-in-class AI translation and dubbing, unmatched control over AI dubbing, and seamless workflow integration. Dubformer caters to media companies, localization companies, and businesses aiming to reach international audiences.

AI dubbing platform for expressive text-to-speech and professional voice-over production
Yueyin AI Dubbing is an AI-powered text-to-speech and voice-over platform developed by Zhipianbang. It converts text into natural-sounding speech using emotionally expressive AI voices and supports professional human voice-over services. The platform offers nearly 1,000 voices across languages, accents, genders, ages, and industry scenarios. It supports single-speaker and multi-speaker dubbing, pronunciation controls for polyphonic characters, numbers, dates, amounts, decimals, pauses, and telephone numbers, as well as WAV audio, SRT subtitle generation, online music, voice cloning, commercial licensing, and related video tools.

AI-powered audio processing platform for voice cloning, noise reduction, and audio translation.
AudioPod AI is an advanced AI-powered platform for noise reduction, voice cloning, and audio processing. It allows users to remove background noise, clone voices, separate speakers, translate audio, and more. It's designed for podcasts, content localization, and professional audio production.

AI-powered text to voice generator with realistic voices for creators and enterprises.
PlayAI is an AI-powered text to voice generator that converts text into realistic Text to Speech (TTS) audio using online AI Voice Generator and synthetic voices. It allows users to instantly convert text into natural-sounding speech and download it as MP3 and WAV audio files. PlayAI offers a platform for creators and enterprises with a low latency Text to Speech API and a library of 200+ realistic AI voices.

AI voice generator with 1000+ voices in 140 languages and voice cloning.
Listnr AI is an award-winning AI Voice Generator and text to speech software with 1000+ voices in 140 languages. It offers realistic AI Voices with an Online Text to Video Generator and voice cloning capabilities. Trusted by 2,500,000+ users, Listnr AI is used for creating voiceovers for various content needs, including shorts, TikToks, Reels, YouTube videos, gaming, podcasts, sales and social media, and audiobooks.

AI voice generator with voice cloning for text-to-speech and speech-to-speech conversion.
Resemble AI is an end-to-end AI voice toolbox engineered for enterprises prioritizing safety and security. It offers AI voice generation with voice cloning for text to speech and speech to speech. Users can clone their voice for free with Resemble's realistic AI voice generator and create voices using real-time speech to speech and text to speech.

Free AI tool for video dubbing, translation, and voice generation.
AI Dubbing is a free online video dubbing tool that translates videos into multiple languages, generates speech based on scripts, and achieves accurate lip sync using advanced AI technology. It provides natural and smooth high-quality dubbing services, supporting over 20 languages and 100+ tones. The platform also offers text-to-speech conversion and voice cloning technology, allowing users to create their own exclusive AI voice from recording samples. It's designed for creators, educators, and businesses to streamline their content localization and production processes.

Musicfy AI: Create AI voice clones, convert voices, and isolate song tracks for music creation.
Musicfy is an AI music assistant that allows users to create AI clones of their voices and use them in any song. It offers features like AI voice conversion, stem splitters, and the ability to create AI covers using various character voices. Musicfy aims to empower artists and creators in the new era of technology.

AI platform for voice cloning, custom synthesis, and multimedia face swapping.
VoidMagic is an AI-powered voice cloning and custom voice synthesis platform that allows users to create high-quality vocal content. It specializes in replicating celebrity voices, cloning unique individual voices from samples, and generating entirely new vocal identities from imagination. Beyond audio, the platform also offers AI face-swapping capabilities for images, videos, and GIFs, providing a comprehensive suite for digital identity transformation.

kikivoice is a free AI voice cloning platform—no sign-up, 75+ languages, up to 99% similarity, with 3 built-in models (Core/Pro/Multilingual).
KikiVoice is a free, creator-first AI voice cloning + TTS platform. Upload a short voice sample to create a realistic voice clone, then generate natural voiceovers from text in minutes—no registration required. It’s built for high timbre fidelity (up to 99% similarity in many cases), fast iteration, and multilingual creation.
Key features
- Voice cloning from a few seconds of audio
- Text-to-speech generation with your cloned voice
- 75+ languages with multilingual & cross-lingual voice creation
- Accent options (model-dependent) for more natural localization
- Quick generation in minutes for rapid content workflows
Built-in models
- Kiki Core: Speed + stability for quick drafts and consistent output
- Kiki Pro: Emotion expression + advanced controls for higher-quality, pro results
- Kiki Multilingual: 75+ languages and accent support for global content scaling
Creator workflow
Upload audio → paste your script → pick a model → generate and download. Seamlessly switch between Core/Pro/Multilingual to balance speed, quality, emotion, and language coverage.

Free real-time AI voice changer with voice cloning and custom integration.
Voice.ai is a free real-time AI voice changer that offers features like voice cloning and custom voice integration in apps. It's designed for streamers, gamers, and businesses for meetings and calls. The platform boasts a decentralized UGC platform for voices and supports various apps and platforms. Users can modify their voice, select from the Voice Universe, or clone any voice they want.

AI platform for subtitles, translation, and dubbing.
Checksub is an AI-powered platform that automatically generates subtitles, translates videos into over 200 languages, and dubs videos with realistic AI voices. It offers voice cloning, lip-syncing, and advanced online editing to maximize the impact of videos for training, social media, and audience growth.

Open-source fine-tuning & reinforcement learning for LLMs. 🦥
Unsloth makes it super easy for you to train text-to-speech (TTS), diffusion, multimodal/image and text models like Llama 3 100% locally or for free on platforms such as Google Colab and Kaggle.
We streamline the entire training workflow, including model loading, quantizing, training, evaluating, running, saving, exporting, and integrations with inference engines like Ollama, llama.cpp, and vLLM.

Kits AI provides studio-quality AI music tools for producers, including voice cloning and mastering.
Kits AI provides studio-quality AI music tools designed to streamline and improve producer workflows. It offers AI voice cloning, the ability to sing like anyone, play any instrument, vocal isolation, stem separation, and AI mastering, all 100% royalty-free. Kits AI is committed to the responsible and ethical use of AI in vocal technology, ensuring every voice used in their models is ethically licensed and securely sourced directly from the artists themselves, with fair artist compensation through revenue-sharing models.

AI video platform transforming URLs into engaging video ads with AI avatars.
Jogg.ai is an AI-powered video platform that transforms URLs into engaging video ads in minutes using rich templates and AI avatars. Effortlessly drive more traffic to your website and boost sales. Create your own Avatar or use 240+ ultra realistic AI Avatars from JoggAI to generate engaging UGC video ads.

AI video creation platform turning photos and text into lifelike videos.
VisionStory AI is an AI video creation platform that allows users to create lifelike AI videos from photos and text. It offers features like emotion control, voice cloning, green screen effects, and multilingual support. It caters to video creators, SME marketing, service & agencies, media & entertainment, and learning & development.

Uncensored AI video and avatar generator for creators and developers.
Makefun is an uncensored, all-in-one AI video and image generation toolset designed for creators and developers. It features advanced capabilities including ultra-realistic image-to-video synthesis, AI head and face swapping, talking photos, voice cloning, and text-to-image generation. Powered by state-of-the-art models like Wan 2.6, Veo 3.1, Kling, and Flux.2, it enables the creation of professional videos without cameras, microphones, or physical actors. The platform also includes comprehensive developer-friendly options such as custom AI Avatar APIs and an MCP (Model Context Protocol) server to seamlessly integrate video and avatar features into external applications.

Task automation platform to deploy AI agents with a single click.
TaskAGI is a task automation platform that allows users to deploy swarms of AI agents with a single click. It provides tools to build, test, and deploy powerful AI agents using an intuitive workflow builder. Users can visually design and configure agents, integrate with various platforms, and monitor performance. TaskAGI also offers pre-built popular agents for customer support, personalized outreach, and meme marketing.

Audimee is a voice-to-voice tool for transforming vocals with studio-quality models.
Audimee is a voice-to-voice tool that lets you transform any vocal using studio-quality models. It allows users to convert vocals with royalty-free voices, train their own voices, and create copyright-free cover vocals. Audimee offers features like vocal conversion, voice training, vocal isolation, voice mixing, and harmony creation.

AI voice solution for content creation with text-to-speech, dubbing, and voice cloning.
Vbee AIVoice is an AI-powered voice solution designed for content creators. It leverages advanced speech technologies like speech synthesis, translation, and recognition to enable the creation of engaging and effective content. It offers features like text-to-speech, AI dubbing, and voice cloning, catering to various content creation needs.

Voice AI platform with ultra-realistic voice solutions for developers and interactive voice apps.
Cartesia is a voice AI platform that offers ultra-realistic voice AI solutions. It provides developers with tools for real-time AI voices, voice cloning, and voice infilling. Cartesia's Sonic model delivers low-latency, high-quality voice AI for interactive voice apps, suitable for real-time voice agents with best-in-class pronunciations. It supports seamless integrations with platforms like Twilio, Pipecat, LiveKit, and Rasa, and offers native speech in 15 languages. Cartesia aims to build the next generation of AI: ubiquitous, interactive intelligence that runs wherever you are.

AI voice generator and content creation tool with realistic AI voices and avatars.
Typecast API is a text-to-speech API designed for developers building conversational AI, content automation pipelines, and voice-enabled applications.
Built on SSFM v3.0 (Speech Synthesis Foundation Model), it automatically reads emotional context from text and delivers the right tone — no manual tagging required. Developers get
700+ expressive AI voices across 38 languages, with support for real-time streaming, batch processing, and webhook-based async flows.
Key reasons teams choose Typecast over alternatives:
• 700+ expressive AI voices: diverse characters across age, gender, and personality —
ready for any product persona, NPC, companion, or narrator
• Smart Emotion: automatically reads text context and delivers the right tone,
no manual tagging
• Real-time streaming API: optimized for conversational AI with no latency gaps
• QuickClone: create a custom branded voice from just 5+ seconds of audio
• Accessible pricing: free tier with 30,000 credits/month, no credit card required
Production references:
• Streaming platforms — real-time TTS serving tens of thousands of concurrent users
with zero latency
• Game studios — NPC voice integration via API across titles
• Content automation — hundreds of short-form videos produced daily via n8n pipelines
• AI companion apps — 6x engagement lift vs. non-voiced interactions
Hot Articles
How to Build a Multimodal AI Knowledge Base With Gemini Embedding 2
NASA 和 IBM 开源月球模型
Introducing the Australian Youth Safety Blueprint
GLM-5.3-FlashX 上线,智谱把国产卡上的推理速度顶到 200 token/s
Google’s Gemini is the latest AI model to hack other companies
How to Choose Hardware for Running Local LLMs, and Know Exactly When It Beats the Claude API
Latest Articles
Google’s Gemini is the latest AI model to hack other companies
How to Choose Hardware for Running Local LLMs, and Know Exactly When It Beats the Claude API
How to Build a Multimodal AI Knowledge Base With Gemini Embedding 2
Announcing Grok-1.5
Google DeepMind launches institute to widen the AGI debate
Google’s new ‘CC’ is an AI agent that helps families run their households
Hot Tags