Voice Generation & Conversion 3007
All Categories
AI Celebrity Voice Generator 34
AI Dubbing 129
AI Podcast 123
AI Podcast Clip Generator 25
AI Podcast Editing 18
AI Recording 52
AI Speech Recognition 151
AI Speech Synthesis 104
AI Speech-to-Text 350
AI Text-to-Speech 385
AI Transcriber 158
AI Transcription 391
AI Voice Assistants 171
AI Voice Changer 58
AI Voice Cloning 213
AI Voice Enhancer 30
AI Voice Generator 343
AI Voice Over 144
Audio To Text AI 121
Tiktok AI Voice Generator 7

Free and premium stock videos, music, and AI tools.
Coverr is a platform offering a vast library of free and premium HD and 4K stock video footage, royalty-free music, and a suite of AI-powered creative tools. These tools include an AI Video Generator, AI Images Generator, AI Voice Over, and AI Sound Effects. It's designed for both personal and commercial use, providing high-quality digital assets and innovative AI capabilities to enhance various creative projects.

Arcade is a platform for creating interactive product demos for various teams and purposes.
Arcade is an interactive demo platform that allows teams to create effortlessly beautiful demos in minutes. It provides tools for marketing, product, sales, customer success, enablement, and training teams to build compelling, on-brand demos to drive leads, boost product adoption, accelerate sales cycles, educate customers, and improve training content. Arcade integrates with other tools and offers features like browser extension capture, desktop app capture, Figma plugin, chapters, call-to-action buttons, export to GIF/Video, white-labeled Arcades, forms, product analytics, integrations, branching, custom links, camera recording, custom variables, synthetic voiceover, hotspots, and callouts.

AI-powered platform for creating interactive product demos to improve communication and engagement.
Supademo helps teams communicate products more effectively using beautiful, AI-powered interactive demos. 4000+ members in customer success, product and marketing embed Supademo across support docs, onboarding, and websites to drive adoption and engagement. Create engaging interactive product demos with AI. Trusted by 50k+ businesses to solve their customer challenges.

Voice AI platform for transcription, voice agents, and speech processing
Smallest AI is a voice AI platform offering speech-to-text, text-to-speech, speech-to-speech, voice cloning, and real-time voice agent technologies. Its Pulse speech-to-text models provide accurate transcription across 38+ languages, global accents, and dialects with latency as low as 64 milliseconds. The platform also supports speaker diarization, sentiment and emotion recognition, language identification, voice agent orchestration, telephony, knowledge bases, and enterprise deployment.

AI-powered audio and video transcription service with summarization and collaboration features.
SoundType AI is an AI-powered audio and video transcription service that converts audio and video files into searchable text. It offers features such as speaker recognition, AI summarization, and interactive chat with audio content. It is designed to improve productivity by integrating transcription, editing, summarization, and collaboration into a single workflow.

Transcribe Any Audio to Text with Audio Transcriber AI Free Online
Audio Transcriber AI is a free online tool designed to convert audio files into text quickly and accurately. It supports multiple audio formats and lets you transcribe without installing or relying on any additional software.

Instantly convert your audio into accurate, searchable text with world-class AI.
Audioconvert.ai is a free, AI-powered tool that converts audio to accurate text in minutes, offering high-quality transcription with speaker detection.

Transcribe audio, video, YouTube links, and recordings into accurate, searchable text.
Audio Converter AI makes it easy to transcribe audio to text online without installing software. Convert interviews, meetings, lectures, podcasts, videos, voice recordings, and YouTube content into accurate, searchable, and editable transcripts.
The AI transcription engine supports more than 200 languages and can automatically detect the source language. Speaker recognition helps separate conversations, while timestamps make it easy to find important moments in long recordings. After transcription, you can use AI-generated summaries and smart notes to understand content faster and create reusable materials.
Audio Converter AI supports popular formats including MP3, MP4, M4A, WAV, WEBM, MOV, AIFF, OPUS, FLAC, AVI, MKV, FLV, and 3GPP. Files can be up to 10GB, with multiple tasks supported in the queue. Your content is processed in encrypted environments, and files are automatically removed after processing.

AI face swap tool for videos and photos with various AI features.
Swapfaces AI is a simple and easy-to-use online AI face swap tool that provides a brand new face-swapping experience for videos and photos. It leverages advanced AI algorithms for seamless and natural results, allowing users to unleash their creativity by changing faces in various media. The platform supports a wide range of AI tools beyond just face swapping, including clothes swapping, image and video enhancement, hair swapping, and more.

AI-powered language learning and assessment platform with AI tutors and assessments in 60+ languages.
Hallo is an AI-powered language learning platform for speaking. It provides fast, affordable, and accurate AI-driven language assessments across speaking, writing, listening, and reading skills, available in over 60 languages. It also offers AI Language Tutor.

AI tool for summarizing, translating, and extracting information from PDFs and other file types.
Coral AI is an AI-powered tool that helps users summarize, find information, translate, and get citations from PDF documents in seconds. It works in over 90 languages and is trusted by researchers and professionals. It can also be used to summarize YouTube videos, transcribe audio, and summarize PowerPoints.

AI-powered news platform providing summaries, insights, and interactive podcasts.
NewsBang is an AI-powered news platform that delivers smarter news and deeper insights behind every headline. It offers concise news summaries, interactive AI podcasts, and the ability to ask AI about any topic. NewsBang aims to provide unbiased news and limitless insights using Generative AI.

Weekly AI newsletter and podcast for techno-optimists and innovators.
Lore Brief is a weekly newsletter and podcast that provides insights into AI breakthroughs, explains why they matter, and offers proof that things are getting better. It caters to techno-optimists and innovators, delivering actionable playbooks, live tool leaderboards, and expert analysis every Friday. The platform aims to help readers master AI in just 5 minutes a week, keeping them informed about the latest developments in the field.

ICAI: AI innovation hub connecting academia, industry, and government in the Netherlands.
ICAI, the Innovation Center for Artificial Intelligence, brings together knowledge institutes, industry, and governmental and societal partners in the Netherlands to develop talent and technology in the area of artificial intelligence. It offers PHD positions, expert collaborations, a launchpad for ecosystem support, and organizes public events to share knowledge and experiences.

API-based voice and SMS platform for communication, data capture, AI analysis, and CRM integration.
Callr is an API-based voice and SMS platform that seamlessly integrates into your product, enabling powerful communication, data capture, AI analysis, and CRM integration. It allows turning conversations into actionable data.

AI-powered tool for accent identification and speech analysis.
Accent Guesser is an AI-powered tool designed for speech analysis, focusing on identifying and analyzing accents. It utilizes deep learning to analyze voice patterns, providing quick and reliable accent analysis. The platform aims to offer insights into users' linguistic backgrounds and enhance communication skills through accent identification and analysis. It is designed with a user-centric interface for ease of use and offers features like global accent recognition and comprehensive data analysis to improve accuracy.

Babbly is an AI-powered tool for early speech therapy and infant development monitoring.
Babbly is an early speech therapy tool that transforms playtime into progress. It uses AI-powered infant speech and brain development monitoring to identify the risk of developmental delays as early as 9 months. Babbly helps parents understand their child’s development by analyzing and monitoring their language progression and recommending activities to accelerate their development. It provides objective data to inform parental intuition and helps parents find out if their child is at risk of speech and language delays, which can be a sign of developmental conditions such as autism.

Kardome offers voice user interface technology for clear voice command input in any environment.
Kardome’s voice user interface technology clusters speech signals based on location, giving clear real-time voice command input and audio output in any environment. Kardome’s AI technology offers an all-in-one solution for manufacturers and OEMs looking to improve their existing speech recognition systems. Kardome’s break through technology improves voice recognition accuracy in challenging soundscapes, transforming voice UI from a cloud-dependent experience to a secure, real-time, and customizable user experience driven by neural network technology that is deployable to any smart device.

Speech recognition and translation software for real-time typing, transcription, and subtitle generation.
SpeechPulse is a speech recognition and translation software that uses your computer’s microphone for real-time speech recognition. It can type into your favorite apps, including text editors, web browsers, and office applications. It can also transcribe audio/video files and generate subtitles. It supports offline speech recognition for ultimate privacy and transcription in 99 languages, including English translation.

AI solutions for audio analysis and speech emotion recognition, enabling empathetic AI interactions.
audEERING provides advanced AI solutions for audio analysis and speech emotion recognition. Their technology transforms industries by enabling machines to understand and respond to human vocal expression, creating empathetic AI interactions. They offer products like devAIce®, devAIce® XR, and AI SoundLab, catering to various use cases such as market research, automotive, robotics, healthcare, and extended reality applications.

AI-powered TOEFL Speaking prep with SpeechRater™ for accurate feedback and score prediction.
My Speaking Score is an AI-powered platform designed to help non-native English speakers prepare for the TOEFL Speaking section. It utilizes ETS's SpeechRater™ technology to provide accurate score predictions and actionable feedback on response delivery, language use, and topic development. The platform offers unlimited practice tests, sharable reports, and personalized insights to help users improve their speaking performance and achieve their target TOEFL score.

AI-powered tool to automatically remove profanity from videos.
Bleepify is an AI-powered tool that automatically removes profanity from videos. It uses advanced AI models to detect and censor offensive language from audio and video files with speed and precision. It supports over 40 languages and allows users to edit, review, and download videos effortlessly. Bleepify is designed for podcasters, video creators, and media managers to ensure their content is clean, professional, and audience-ready.

AI-powered English speaking coach for employees, offering personalized feedback and secure language training.
Lucida AI is an AI-powered English speaking coach designed to help employees enhance their communication skills. It offers personalized feedback on pronunciation, grammar, vocabulary, and fluency through real-time conversations with Lucy, an AI coach. Lucida AI prioritizes privacy with end-to-end encryption and can be tailored to company regulations. It provides comprehensive language training at an affordable price, ensuring every team member can benefit from advanced AI-driven coaching.

Accurate speech-to-text API and speech recognition service with various features and language support.
Rev AI is a speech-to-text API and speech recognition service that offers accurate transcription at 0.3¢/min. It provides asynchronous and streaming APIs, human transcription services, and insights like topic extraction and sentiment analysis. Rev AI supports multiple languages and offers features like language identification and forced alignment.
Hot Articles
How to Build a Multimodal AI Knowledge Base With Gemini Embedding 2
NASA 和 IBM 开源月球模型
Introducing the Australian Youth Safety Blueprint
GLM-5.3-FlashX 上线,智谱把国产卡上的推理速度顶到 200 token/s
Google’s Gemini is the latest AI model to hack other companies
How to Choose Hardware for Running Local LLMs, and Know Exactly When It Beats the Claude API
Latest Articles
Google’s Gemini is the latest AI model to hack other companies
How to Choose Hardware for Running Local LLMs, and Know Exactly When It Beats the Claude API
How to Build a Multimodal AI Knowledge Base With Gemini Embedding 2
Announcing Grok-1.5
Google DeepMind launches institute to widen the AGI debate
Google’s new ‘CC’ is an AI agent that helps families run their households
Hot Tags