Voice Generation & Conversion 3170

Converts online meeting recordings into branded, short-form video clips automatically. ProdShort is an AI-powered platform designed to transform online meeting recordings into engaging short-form video clips. Instead of generating synthetic content, it captures real value from existing meetings on platforms like Zoom, Google Meet, and Microsoft Teams. It automatically records sessions, identifies key moments, and allows users to refine them into polished shorts for LinkedIn, Twitter, and TikTok. The tool features TikTok-style animated captions, custom branding with logos, and professional templates to ensure brand consistency without the need for manual editing or scripting.
AI-powered tool to convert unstructured text to structured JSON data without REGEX. JSON Scout is a tool that converts unstructured text content into structured JSON data using AI, eliminating the need for REGEX. It helps users fetch insights from their data by defining the desired output, inputting content, and processing it to return structured data.
A secure, EU-based meeting platform with AI tools and strong privacy features. Eyre is a compliance-first meeting platform that prioritizes privacy and security. It offers AI-powered meeting agendas, transcripts, summaries, action items, and task management, suitable for work, learning, and lifestyle projects. Eyre aims to transform mundane meetings into engaging and interactive sessions by providing private, encrypted EU-based meetings with AI-powered summaries, tasks, and transcripts, adhering to European privacy standards.
Voice-powered task management platform for efficient project and note organization. Muchtodo is a task management platform that leverages voice input and artificial intelligence to help users manage projects, tasks, and notes more efficiently. It allows users to create and organize their work using speech-to-text technology, offering a hands-free approach to productivity. Muchtodo supports multiple languages and provides features like project boards and calendar integration to streamline task management.
Voice-based productivity app for creating tasks and events. Intellisay is a productivity app that uses voice input to create a list of tasks and events, helping users optimize their day. By speaking about their day for two minutes, users receive an actionable list to achieve success.
PixtaAI offers data annotation, collection, and licensing for AI/ML/CV projects. PixtaAI provides high-quality data annotation and data collection services for AI, Machine Learning, and Computer Vision projects at any scale. It offers a licensing platform to get trusted, order-made visual datasets easily and integrate them into machine learning pipelines seamlessly. PixtaAI empowers users with reliable and accurate data through its technology and right management system, providing access to high-quality, licensed data from multiple data categories.
AI-powered digital asset management solution for managing and accessing media assets. Evolphin Zoom MAM is a cloud-based, AI-powered digital/production/media asset management solution. It manages the total lifecycle of all digital assets, including hard-to-manage video files, making them accessible to remote and on-premises teams. It simplifies image, audio, and video workflows for Creative, Marketing, and IT teams.
Recapio is an AI second brain for capturing, organizing, and transforming knowledge. Recapio is an AI-powered second brain that helps users capture, organize, and transform knowledge from various sources like YouTube videos, web articles, and more. It offers features like AI summarization, chat with video, and knowledge organization to empower learning and productivity.
AI-powered platform for unified digital workflows and enhanced knowledge management. Unifie by Typeless is a platform designed to transform digital workflows, reduce cognitive load, and enhance productivity by unifying digital processes. It aims to supercharge your knowledge journey with AI, allowing users to create, organize, and discover information efficiently. It offers features like seamless research, integration of personal documents, uninterrupted thought flow, and intuitive note-taking.
All-in-one AI platform for professional video, image, music, and voiceover creation. Artta AI is an all-in-one AI creative platform designed for generating professional videos, images, music, and voiceovers. It integrates multiple leading AI models such as Sora 2, Veo 3, Flux, DALL-E, Midjourney, Stable Diffusion, and Kling AI, enabling creators to transform ideas into high-quality content faster. The platform offers automated workflows, professional asset management, advanced effects, character consistency, real-time collaboration, and API integrations, catering to a wide range of users from social media content creators to filmmakers.
AI text-to-speech tool with 300+ voices and voice cloning. DupDub is an AI-enabled text-to-speech tool based on an industry-leading in-house speech synthesis system. It supports 300+ AI voices with different emotions and provides professional voice cloning services. It also offers AI tools for voiceovers, dubbing, avatars, and writing.
Netwrck is an AI character marketplace with chat, voice, and art generation features. Netwrck is an AI Character Marketplace where you can create AI Characters and earn NETW tokens by engaging the community. Chat to your favorite AI Characters and socialize. Netwrck also offers AI Chat, AI Characters, AI Voice Chat, AI Art Generator and AI Chatbots.
Aquin Lucid is an AI-powered browser focused on productivity and ease of use. Aquin Lucid is a browser developed by Team Aquin, designed to enhance user productivity and ease of use. It aims to create a new genre of applications by combining AI, productivity tools, and universal accessibility. It features an AI search engine, Zen Mode for focused browsing, and tools for interacting with web content in new ways, such as talking to webpages and searches.
Proactive AI assistant for sales professionals to automate emails, scheduling, and deal prioritization. Demi AI is a proactive AI assistant specifically designed for client-facing professionals like sales reps and account executives. It integrates directly with Gmail and Outlook to prioritize revenue-driving emails using dynamic labels, draft replies in the user's unique voice, and automate follow-ups. Beyond email, it assists with scheduling meetings and enhancing meeting transcriptions into actionable follow-up tasks, helping professionals close more deals while spending less time managing their inboxes.
Conversational AI platform for enterprise customer and employee support automation. Enterprise Bot provides conversational automation solutions for enterprises, leveraging LLMs and ChatGPT to transform chat, email, and voice interactions. Their platform offers AI-powered chatbots, email response automation, and voice bots designed to improve customer and employee support.
AI-powered transcription and content generation service with high accuracy and multiple features. WhisperTranscribe is an AI-powered transcription service that converts audio and video into accurate transcripts with timestamps. It allows users to generate new content from transcripts, such as summaries, blog posts, and social media posts, using GPT prompts. It supports 55+ languages, speaker recognition, and flexible export options. No subscription is required, and a free trial is available.
Uberduck is an AI platform for voice-over, text-to-speech, voice cloning, and AI music generation. Uberduck provides users with tools to create voice-over audio with over 5,000 expressive voices, custom voice clones, APIs to build audio applications, and AI-generated raps. It also offers features like text to speech, voice conversion, and AI music generation. There is a case study to demonstrate how it can be used to create personalized media and a waitlist to join the upcoming Uberbots platform.
AI-powered platform for video editing, text-to-speech, voice cloning, and localization. Wavel AI is an AI-powered platform that offers a suite of tools for video editing, text-to-speech conversion, voice cloning, translation, and more. It aims to simplify video creation and localization, making it accessible to users with varying levels of expertise. The platform provides solutions for AI dubbing, video editing, voice cloning, subtitle generation, and other video-related tasks.
AI-powered transcription service converting audio and video to text with high accuracy and useful features. AudioScribe.io is an AI-powered transcription tool designed to convert audio and video recordings into accurate, readable text transcripts with ease. It offers features like automated meeting join & record, support for audio and video files, full-text search, AI summaries, and speaker identification. It caters to freelancers, small businesses, and large enterprises, ensuring no word is missed in meetings, interviews, or important conversations.
Multilingual voice AI routing and infrastructure platform Speko is a voice AI routing and infrastructure platform that provides a unified API for speech-to-text, language models, and text-to-speech. It benchmarks speech and language models across multiple languages, routes each session to suitable-performing models, and supports provider-direct or managed voice workflows through integrations with LiveKit, Pipecat, OpenAPI, AsyncAPI, and MCP.
Scribewave is an accurate online speech-to-text tool with transcription, translation, and subtitling in 90+ languages. Scribewave is the most accurate online speech-to-text tool for all audio and video files. It offers subtitles, translations, and transcripts in 90+ languages. Features include flawless transcription, 100% privacy, automatic subtitles, automatic captions, transcription translation, transcripts, speech-to-text, and audio to text conversion.
Ultra-low-latency voice AI APIs for speech generation, transcription, translation, and cloning Gradium is a voice AI platform for developers that provides ultra-low-latency text-to-speech, speech-to-text, speech-to-speech translation, live translation, voice cloning, and on-device text-to-speech through a unified API. It is designed for building real-time voice agents and conversational applications, with expressive speech generation, accurate transcription, multilingual support, speaker cloning, bidirectional WebSocket streaming, scalable concurrency, and deployment options including cloud, dedicated instances, self-hosted, and on-premises infrastructure.
Write with your voice on any website, with 99% accuracy, in over 90 languages. BlabbyAI is a smart voice-to-text tool that lets you write with your voice on any website. Powered by OpenAI’s Whisper model, it turns your speech into text quickly and with high accuracy. Beyond simple transcription, BlabbyAI offers Custom Modes that shape your speech into exactly what you need. Grammar Mode fixes and polishes your sentences, like Grammarly. Email Mode transforms what you say into professional, ready-to-send emails. Custom Mode lets you create your own with a simple prompt, telling the AI how to format or improve your words. BlabbyAI is perfect for professionals, writers, and anyone who wants to get more done just by talking.
Opensource voice-to-text macOS app with local AI for privacy and offline use. VoiceInk is an opensource voice-to-text app for macOS that transcribes what you say to text almost instantly with near-perfect accuracy. It uses local AI models to transcribe your speech to text, enabling offline functionality and ensuring data privacy. All data is stored locally, with optional AI enhancement.