Voice Generation & Conversion 3007
All Categories
AI Celebrity Voice Generator 34
AI Dubbing 129
AI Podcast 123
AI Podcast Clip Generator 25
AI Podcast Editing 18
AI Recording 52
AI Speech Recognition 151
AI Speech Synthesis 104
AI Speech-to-Text 350
AI Text-to-Speech 385
AI Transcriber 158
AI Transcription 391
AI Voice Assistants 171
AI Voice Changer 58
AI Voice Cloning 213
AI Voice Enhancer 30
AI Voice Generator 343
AI Voice Over 144
Audio To Text AI 121
Tiktok AI Voice Generator 7

Open-source Slack bot for summarizing content and voice communication.
myGPTReader is an open-source Slack bot that summarizes webpages, ebooks, and YouTube videos, and communicates via voice. It allows users to read and chat with AI bots within Slack, offering features like web reading, document reading, voice chat, and hot news summaries.

AI tool that decodes messy handwritten medical prescriptions into clear, readable, and structured text.
Doctor Handwriting Reader AI is a specialized tool designed to decode and interpret messy, handwritten medical prescriptions and notes. Using advanced AI-powered OCR, it converts difficult-to-read clinical handwriting into clear, structured text. The platform not only extracts raw text but also organizes information into key medical categories like medication names, dosages, possible conditions, and next steps, providing a plain-English summary for better understanding by patients and healthcare professionals.

AI reader and roleplay app for branching fiction
Foreverse Xinmeng is a mobile AI reader and tavern roleplay app for long fiction, branching continuations, character cards, worldbooks, and BYOK model control.

Audioread converts text to audio for listening in podcast apps using AI voices.
Audioread turns articles, PDFs, emails, and RSS feeds into audio, allowing users to listen in their podcast player. It uses ultra-realistic AI voices to read text aloud, enabling users to listen while exercising, cooking, or commuting. Audioread generates a private podcast RSS feed that can be subscribed to in any podcast app, such as Apple Podcasts, Google Podcasts, and Spotify.

AI reader that turns PDFs, ebooks, and webpages into natural-sounding speech.
Readio is an AI-powered text-to-speech application designed to transform various text formats—including webpages, PDFs, EPUBs, and documents—into natural-sounding speech. It utilizes advanced neural networks, specifically OpenAI TTS voices, to provide lifelike intonation and clear audio across over 140 languages and accents. The platform is built for long-form listening, allowing users to consume content hands-free while maintaining focus through features like synchronized live highlighting and automatic scrolling.

Free AI text-to-speech tool using ChatGPT voices for an immersive listening experience.
GPT Reader is a free AI text-to-speech (TTS) tool that utilizes ChatGPT's premium voices to provide an unparalleled listening experience. It allows users to convert text from PDFs, articles, and documents into natural-sounding speech. With features like dark/light mode, adjustable playback speeds, pause and resume functionality, and a full-screen UI, GPT Reader offers an immersive way to engage with content.

Text-to-speech app for PDFs, Word docs, and more to boost productivity.
Audeus is an immersive text-to-speech (TTS) reader app designed for PDFs, Word documents, and more. It allows users to listen to their documents, text, and web articles to save time and boost productivity. Audeus supports various file formats, including PDF, Word (docx), non-DRM EPUB, and integrates with Chrome, Canva, and OpenAI. It offers lifelike voices, synced text highlighting, and customizable playback speeds to enhance comprehension and retention.

TourMe provides curated travel content for personalized, gamified learning experiences worldwide.
TourMe is a platform designed to enhance travel experiences by providing curated content in multiple languages. It allows travelers to quickly learn about landmarks and locations, offering personalized, gamified learning experiences. TourMe aims to transform travelers into both learners and guides, enabling them to explore the world's stories at their own pace, language, and style.

XspaceGPT converts Twitter Spaces to text with AI summaries and multi-language support.
XspaceGPT transforms Twitter Spaces into text with summaries, outlines, highlights, and multi-language support. It helps users discover top live spaces and influential hosts, download spaces, and explore a library to expand their knowledge. It offers AI-generated summaries and mind maps, converting audio to text effortlessly.
AI video editing software for YouTube creators to optimize workflow and content.
Gling is an AI-powered video editing software designed for YouTube creators. It streamlines the editing process by automatically cutting out bad takes, silent moments, filler words, and background noise. Gling also offers features like AI captions, auto framing, title generation, and next video suggestions to maximize YouTube success.

All-in-one AI anime studio for scripts, storyboards, and dubbed videos.
MkAnime AI is an all-in-one AI anime production studio designed for anime and manga enthusiasts, solo creators, and professional teams. It streamlines the entire creative process by transforming a single story prompt into structured outlines, scripts, character sheets, and storyboards. The platform features tools for character consistency, AI dubbing with lip-sync, and multi-format video export in 9:16 and 16:9 ratios. It allows users to manage complex series with recurring casts and high-quality motion comic assets within a single browser-based workflow.

AudioBook Bot uses AI to convert text to audiobooks with multiple voices.
AudioBook Bot converts written works to audio works using text to speech AI. Using AudioBook Bot, you can create character rich audiobooks using multiple licensed voices or read books and podcasts in your own voice. It is a one-click Audiobook creation software that uses generative AI to convert text to speech. With some additional annotations, it can provide your book with a whole cast of characters. You can also narrate the book in your own voice with a minimal sample.

AI tool for quick, professional video voiceovers using GPT-4o and ElevenLabs.
Overvoice is a tool that combines GPT-4o and ElevenLabs to automatically add professional-grade voiceovers to videos in under a minute. It eliminates the need for manual script writing and tedious voiceovers, streamlining the voiceover creation process to enhance conversion rates with demo videos.

AI-powered platform for generating professional voice-overs for videos.
NarrateVideoAI is an AI-powered platform that transforms videos with AI narration, providing professional voice-overs in multiple languages and styles. It allows users to automatically generate voice-overs for their videos using advanced AI technology, without requiring any technical expertise. The platform supports multiple languages, offers various voice options and styles, and ensures fast processing and high-quality voice synthesis.

Narrai simplifies adding voiceovers to videos with AI-generated scripts and music.
Narrai simplifies adding relevant voiceovers for videos in a simple delightful flow. Whether for personal, social or business, Narrai generates a unique script, voice generation and background music merged for posting or saving. It allows users to transform text or visual media into captivating, spoken narratives, enhancing the overall impact and accessibility of the content.
AI platform to convert text files into human-like voiceovers with voice cloning.
Cugent is an AI-powered platform that turns PDF, Doc, and text files into MP3 voiceovers. It supports different voice types in all supported languages and offers multilingual voice types for German, Dutch, Italian, and French scripts. It also features voice cloning, allowing users to use their own voice in all supported languages.

AI dubbing and voiceover tool for media and entertainment with cost-effective localization.
Dubformer is an AI dubbing and voiceover tool for media and entertainment, offering monetization-ready dubbing that sounds like real humans. It provides cost-effective AI localization solutions, including best-in-class AI translation and dubbing, unmatched control over AI dubbing, and seamless workflow integration. Dubformer caters to media companies, localization companies, and businesses aiming to reach international audiences.

AI dubbing platform for expressive text-to-speech and professional voice-over production
Yueyin AI Dubbing is an AI-powered text-to-speech and voice-over platform developed by Zhipianbang. It converts text into natural-sounding speech using emotionally expressive AI voices and supports professional human voice-over services. The platform offers nearly 1,000 voices across languages, accents, genders, ages, and industry scenarios. It supports single-speaker and multi-speaker dubbing, pronunciation controls for polyphonic characters, numbers, dates, amounts, decimals, pauses, and telephone numbers, as well as WAV audio, SRT subtitle generation, online music, voice cloning, commercial licensing, and related video tools.

Free online AI tool for voice and language transformation.
Voice Changer is a free online AI tool that transforms voices using artificial intelligence technology. It offers a rich library of over 100 AI voices and supports more than 20 different languages, allowing users to easily change their voice or language. It is perfect for creating engaging multilingual audio content, providing natural and realistic voice effects for various applications such as content creation, localization, education, animation, marketing, and development.

All-in-one platform with AI voiceover, transcription, editor, chat, templates, and image generator.
Unmixr is an all-in-one SaaS platform that provides AI Voiceover, AI Transcription, AI Editor, AI Chat, Templates, and AI Image generator. It offers realistic AI voiceovers, accurate AI transcription, and AI-powered content editing and image generation.

AI voice over generator with human-like voices for diverse content creation.
Lazybird is an AI-powered voice over generator that allows users to create human-like automated voice overs for various content types, including videos, podcasts, audiobooks, and educational materials. It offers a wide range of voices, languages, and customization options, aiming to save time and cost in voice over production.

AI-powered audio processing platform for voice cloning, noise reduction, and audio translation.
AudioPod AI is an advanced AI-powered platform for noise reduction, voice cloning, and audio processing. It allows users to remove background noise, clone voices, separate speakers, translate audio, and more. It's designed for podcasts, content localization, and professional audio production.

AI-powered text to voice generator with realistic voices for creators and enterprises.
PlayAI is an AI-powered text to voice generator that converts text into realistic Text to Speech (TTS) audio using online AI Voice Generator and synthetic voices. It allows users to instantly convert text into natural-sounding speech and download it as MP3 and WAV audio files. PlayAI offers a platform for creators and enterprises with a low latency Text to Speech API and a library of 200+ realistic AI voices.

AI tool for generating comics, webtoons, and animations with custom character models.
Autodraft AI is a premier AI-powered tool designed for generating comics, webtoons, and animations. It allows users to train custom character models to achieve character and style consistency. The platform offers advanced tools for voiceovers, character creation, and image-to-animation generation, making it easy to create professional animation videos.
Hot Articles
How to Build a Multimodal AI Knowledge Base With Gemini Embedding 2
NASA 和 IBM 开源月球模型
Introducing the Australian Youth Safety Blueprint
GLM-5.3-FlashX 上线,智谱把国产卡上的推理速度顶到 200 token/s
Google’s Gemini is the latest AI model to hack other companies
How to Choose Hardware for Running Local LLMs, and Know Exactly When It Beats the Claude API
Latest Articles
Google’s Gemini is the latest AI model to hack other companies
How to Choose Hardware for Running Local LLMs, and Know Exactly When It Beats the Claude API
How to Build a Multimodal AI Knowledge Base With Gemini Embedding 2
Announcing Grok-1.5
Google DeepMind launches institute to widen the AGI debate
Google’s new ‘CC’ is an AI agent that helps families run their households
Hot Tags