Text-to-speech 91

Create AI talking head videos from a photo and script in minutes. Lemon Slice (formerly Infinity AI) is a video foundation model that allows you to create expressive, talking characters from a photo and script in minutes. It is ideal for creators, marketers, and businesses. It offers features including resolution, commercial rights, voice cloning, and more, with a free tier available.
Platform for scaling audio content with synthetic voices and publishing tools. BeyondWords is a platform designed to scale audio content production, distribution, and monetization operations. It offers high-quality synthetic voices and audio publishing tools, enabling users to convert text into engaging audio. The platform provides an all-in-one audio CMS and AI voices to enhance the publishing workflow.
FileSpeech converts files to natural speech with multilingual support and offline access. FileSpeech is a platform designed to convert files into natural speech. It supports multiple languages and offers a selection of neural voices. Users can upload files in various formats, including PDFs, EPUBs, and web links, or scan documents using their device's camera. The platform also provides offline features, allowing users to convert and export audio files for listening anywhere.
PollySpeak is a text-to-speech tool for listening to books, documents, and web pages. PollySpeak revolutionizes how we consume content. It allows you to listen to books with lifelike voices, read text from scanned documents, and browse the web with an audio TTS companion. It's an affordable and resilient text-to-speech tool that helps overcome distractions, improve accessibility, and increase reading speed.
ChatTTS: Natural, expressive text-to-speech for dialogue applications in English and Chinese. ChatTTS is a powerful text-to-speech model designed for creating natural and expressive speech, perfect for dialogue-based applications. Supporting both English and Chinese, ChatTTS offers fine-grained control over prosodic features like laughter and pauses.
Natural, real-time voice synthesis for various applications. Advanced Voice from ChatGPT offers natural, real-time voice synthesis with custom instructions, memory, and improved accents. It enables smoother, faster conversations suitable for virtual assistants, audiobooks, customer service, and more. It generates human-like, natural-sounding outputs with real-time processing and high-quality audio output. The system supports interactive dialogue with custom instructions and memory, enhancing conversational speed and smoothness.
Free AI text-to-speech platform for natural-sounding speech conversion. Nemesys Labs offers a free AI-powered text-to-speech platform that instantly converts text into natural-sounding speech. It is designed for content creators, educators, and developers, providing an accessible speech synthesis infrastructure.
AI-generated podcasts from articles, blogs, and news on chosen topics. Nural News pulls the latest articles, blogs, and breaking news on your chosen topic, then turns it into an AI-generated podcast you can listen to anytime.
AI-powered speech generation app for content creators to create engaging audio and video content. Allinpod.ai is a speech generation app for content creators, powered by the latest AI technology, inspired by All-in podcast. It allows users to create engaging audio and video content with AI speech software, enabling them to make All-in podcast Besties talk. It provides prime AI speech and video generation software to create content you've always wanted, discovering the future of podcasting.
AI tool converting articles to podcast-quality audio for effortless listening. Read-this.ai is an AI-powered tool that converts articles into natural, podcast-quality audio with a single click. It allows users to listen to web content effortlessly, transforming the internet into a personal audio library. This service redefines the reading experience by making it accessible and convenient for on-the-go consumption.
Gan AI offers text-to-speech API and an AI playground for experimentation. Gan AI offers the Myna-mini TTS API for research preview, which supports high-fidelity, text-to-speech in all 22 official Indic Languages & English with seamless code-mixing capabilities and free playground access. It also provides an AI Playground where users can experiment with AI avatars, test video personalization, and explore new creative possibilities.
Interactive streaming tool to boost chat engagement with custom TTS, alerts, and more. Tangia is an interactive streaming tool designed to supercharge chat engagement on platforms like Twitch. It offers features like custom TTS, interactions, alerts, media sharing, and a monitor overlay, allowing streamers to create more engaging and interactive experiences for their viewers. Tangia aims to provide streamers with the tools they need to level up their streams and foster a stronger sense of community.
AI content generation platform with tools for text, images, voiceovers, and code. Cannypen is an AI-powered platform designed to generate various types of content, including articles, ads, blog posts, and AI voiceovers. It offers a range of AI tools such as AI Chat Bots, AI Contents, AI Images, AI Voiceovers, AI Speech to Text, and AI Codes. Cannypen aims to help users create content 10X faster with over 70 templates and supports content generation in more than 54 languages.
AI-powered platform offering expert advice and content generation tools. Assistante.app offers 24/7 AI experts for advice in business, health, dating, and more. It's an all-in-one platform for generating AI content and getting advice in minutes. It provides AI chatbots tailored for precise and instant answers, along with tools for image generation, content creation, and document summarization.
Real-time STT/TTS solution using AI-focused Sense Theory for nuanced speech processing. Speech Intellect is the first STT/TTS solution that works in real-time by totally using a new AI-focused mathematical theory — "Sense Theory". It looks at the sense of each word pronounced by the client. It offers speech-to-text, text-to-speech, and combining solutions, leveraging a sense-to-sense algorithm to reproduce text with intonation and tonality. The platform emphasizes security with Amorphous Encryption and provides flexibility in shaping work scenarios for various business needs.
AI voice generator for text-to-speech, cloning, and custom voices. VoiSpark is an AI voice generation platform that enables users to create human-like voices, generate realistic text-to-speech, clone voices, and design custom AI voices. It serves as an all-in-one AI voice toolkit powered by industry-leading AI, offering over 500 natural-sounding AI voices and multi-language support across 30+ languages. The platform is designed for creating studio-quality voiceovers for various content types like videos, podcasts, and apps.
AI platform for instant voice cloning and high-quality multilingual text-to-speech generation. Voiceslab is an AI-powered voice cloning platform designed to create realistic and unique digital replicas of any voice. By analyzing a short audio sample of 10-60 seconds, the technology captures speech patterns, tones, and accents to generate high-quality text-to-speech content. It supports multiple languages and allows users to produce audio content in their own voice or a specific cloned voice without the need for manual recording, making it a valuable tool for content creators and businesses.
AI-powered platform for video editing, text-to-speech, voice cloning, and localization. Wavel AI is an AI-powered platform that offers a suite of tools for video editing, text-to-speech conversion, voice cloning, translation, and more. It aims to simplify video creation and localization, making it accessible to users with varying levels of expertise. The platform provides solutions for AI dubbing, video editing, voice cloning, subtitle generation, and other video-related tasks.
Multilingual voice AI routing and infrastructure platform Speko is a voice AI routing and infrastructure platform that provides a unified API for speech-to-text, language models, and text-to-speech. It benchmarks speech and language models across multiple languages, routes each session to suitable-performing models, and supports provider-direct or managed voice workflows through integrations with LiveKit, Pipecat, OpenAPI, AsyncAPI, and MCP.
Ultra-low-latency voice AI APIs for speech generation, transcription, translation, and cloning Gradium is a voice AI platform for developers that provides ultra-low-latency text-to-speech, speech-to-text, speech-to-speech translation, live translation, voice cloning, and on-device text-to-speech through a unified API. It is designed for building real-time voice agents and conversational applications, with expressive speech generation, accurate transcription, multilingual support, speaker cloning, bidirectional WebSocket streaming, scalable concurrency, and deployment options including cloud, dedicated instances, self-hosted, and on-premises infrastructure.
Novita AI: AI cloud with model APIs, GPU instances, and serverless GPUs. Novita AI provides access to over 100 APIs, including AI image generation and editing with 10,000+ models, and training APIs for custom models. It offers cheap pay-as-you-go services, freeing users from GPU maintenance hassles while building their own products. Novita AI also provides GPU Instances and Serverless GPUs for scaling AI, optimizing performance, and innovating with ease and efficiency.
AI-powered text-to-speech platform with multilingual support and premium audio quality. TxtVoice is a next-generation AI-driven text-to-speech platform that converts text into lifelike voices instantly. It supports over 50 languages, offers real-time conversion, and provides premium audio quality. Users can customize pitch and speed. TxtVoice offers free AI voices and end-to-end encryption to ensure data security and privacy.
All-in-one AI video generator with realistic avatars and text-to-speech. AI STUDIOS by Deepbrain AI is an all-in-one AI video generator that allows users to create videos quickly using simple text prompts. It offers realistic AI avatars, natural text-to-speech capabilities, and powerful AI video editing tools. The platform supports collaboration, video translation, and various templates for different use cases, making video creation accessible to individuals and businesses alike.
AI voice generator with realistic text-to-speech and speech-to-speech capabilities. Respeecher Voice Marketplace is an AI voice generator platform that offers realistic text-to-speech and speech-to-speech capabilities. It provides a range of AI voice solutions for creative and professional projects, including film and TV production, game development, advertising, and more. The platform is trusted by industry leaders and offers high-quality AI voices, including celebrity voices, with a focus on ethical use and legal compliance.