AI Transcriber 157

PodShrink is an AI-powered podcast summarizer that transforms full-length podcast episodes into concise, narrated audio summaries you can listen to on the go. PodShrink lets you pick any podcast episode from a library of thousands of shows, choose your preferred AI voice and summary length (1, 5, or 10 minutes), and get a professionally narrated audio summary in minutes. Unlike text-only summarizers, PodShrink outputs listenable audio — so you consume summaries the same way you'd consume the original podcast. Every episode also includes a full searchable transcript, making it easy to find the exact quote, insight, or moment you need without relistening. Powered by ElevenLabs TTS with 12 premium voices and Google Gemini for intelligent summarization. Free to start, with Standard ($9.99/mo) and Pro ($19.99/mo) plans for power users.
AI-powered podcast production tool for transcripts, show notes, chapters, and clips. Podium is an AI-powered tool designed to streamline podcast production by generating transcripts, show notes, chapters, highlight clips, and more. It aims to save time and enhance content quality for podcasters of all sizes, offering features like AI copywriting for show notes and articles.
All-in-one AI content generation platform to make money. MeduzaAi is an all-in-one platform designed to generate AI content and enable users to start making money in minutes. It offers a range of AI tools, including text, image, code, and chatbot generators, along with speech-to-text and voiceover capabilities. The platform aims to empower users with AI-driven solutions for various content creation needs, from blog posts and social media content to product descriptions and ad copy.
AI platform for YouTube/Vimeo video optimization: chapters, tags, descriptions, titles. VidChapter is an AI-powered platform designed to generate timestamped chapters, optimized tags, descriptions, and engaging titles for YouTube and Vimeo videos. It aims to elevate viewer experience, boost engagement, and improve SEO. The platform offers tools for video optimization, including AI-driven chapter creation, SEO-powered descriptions, captivating titles, smart tag generation, content translation, and transcription.
Create with our AI voice generator, offering text-to-speech, voice cloning, and celebrity/character voices. Access a massive sound effects library and AI story generator for all your creative projects. Discover the ultimate toolkit for audio creation with our comprehensive AI-powered platform. At the core of our service is a state-of-the-art AI Voice Generator that brings your text to life. With a vast selection of hundreds of realistic male, female, and character voices, you can generate high-quality audio for any project. Our Text-to-Speech (TTS) technology is designed for a multitude of applications, from creating engaging voiceovers for videos and podcasts to developing immersive e-learning content and accessibility solutions. Unleash your creativity with our advanced Voice Cloning and Voice Conversion tools. Effortlessly replicate any voice from a simple audio sample, opening up endless possibilities for personalized content. Dive into our extensive library of unique voice styles, including celebrity impersonations, allowing you to have your script read by voices resembling famous figures like Morgan Freeman, Joe Rogan, and even political leaders like Donald Trump and Joe Biden. Our specialized generators cater to every niche, from a Pirate AI Voice Generator and a Horror AI Voice Generator for thematic content to an ASMR AI Voice Generator for relaxation and a Motivational AI Voice Generator for inspiring messages. We also feature a diverse range of character voices from popular culture, such as SpongeBob, Optimus Prime, and various Disney and anime characters. Beyond voice generation, our platform is a complete audio production suite. Explore our massive, ever-expanding Sound Effects Generator, featuring everything from ambient nature sounds and cityscape noises to cartoon boings, cinematic explosions, and futuristic sci-fi effects. This library is an invaluable resource for game developers, filmmakers, and content creators seeking to enhance their projects with professional-grade audio. To further streamline your workflow, we offer a suite of powerful tools. Our AI Voice Recorder allows you to capture audio directly, while the YouTube to MP3 converter makes it easy to extract audio from online videos. For video creators and translators, the AI Timestamp SRT Generator and Timestamped Script-to-Voice Generator provide precise synchronization of audio and text. Additionally, our innovative AI Story Generator can help you craft compelling narratives, providing both inspiration and the tools to voice them. Whether you're producing content for TikTok, creating narrated scary stories, or developing the next hit video game, our platform provides all the tools you need in one place. Experience the future of audio creation with our versatile and user-friendly AI voice and sound generation services.
Your all-in-one creative AI suite for videos, music, voices, avatars, images, and more. Klyra AI is an all-in-one AI platform that empowers creators, marketers, and businesses to produce videos, voiceovers, avatars, images, music, chatbots, blogs, and more using 30+ powerful AI tools — all without technical skills. From text-to-video, image-to-video, voice cloning, and face swapping to advanced photo editing, real-time voice chat, and SEO automation, Klyra combines cutting-edge AI models with an intuitive interface to simplify your creative and productivity workflows. Whether you’re building content, coding, managing social media, or automating blog publishing, Klyra AI helps you work smarter, create faster, and deliver professional results with ease.
Pin voice, photo and video notes onto plans, site photos and 3D models — construction site reporting from your phone. PinMy is a mobile-first visual collaboration app for construction and AECO teams. Open a PDF plan, a site photo, a walkthrough video or a 3D IFC model, tap the exact spot where the issue is, and leave a voice note, photo, video or text right there. Voice notes are transcribed automatically and become searchable — so the person who wasn't on site can still find what was said, and exactly where. Everything captured on site exports as a PDF report in one tap: all of it, or filtered by status, assignee or day. Issues also move in and out of the BIM stack through BCF 2.1 import and export, so PinMy extends the tools the office already runs instead of replacing them. Uploaded IFC models are converted to a lightweight format that actually performs on a phone or tablet. PinMy replaces the chain most site teams still live with — camera roll, WhatsApp groups, paper markups and a spreadsheet typed up at night. Every note becomes a dated, located, attributed record: which defect, on which plan, at which point, said by whom. That matters on the day a dispute starts. No training, no desktop CAD, no BIM license. If someone can send a voice message, they can use PinMy. Free forever for individuals; Premium is €8.99/month + VAT. Available in 8+ languages on iOS, iPadOS, Android, the web and as a Chrome extension. EU-hosted, GDPR-aligned, encrypted in transit and at rest — your content is never used for AI training and never sold.
HaloVoice is an AI-powered real-time voice translator for gaming, streaming, voice chat, and online meetings, featuring voice cloning and bilingual subtitles. HaloVoice is an AI-powered real-time voice translator designed for gaming, live streaming, voice chat, online meetings, and multilingual communication. It translates spoken conversations across languages in real time, helping people communicate naturally without constantly switching to a text translator. Real-Time Voice Translation HaloVoice provides live voice translation while you speak. Instead of typing text into a translator, users can speak naturally and send translated speech directly to other people in real time. Bilingual subtitles display both sides of the conversation, making multilingual communication easier to follow. With AI voice cloning, HaloVoice can preserve characteristics of the speaker’s voice in translated speech, creating a more natural and personal experience. Real-Time Voice Translator for Gaming HaloVoice works as a real-time voice translator for gaming, helping players communicate with international teammates while they play. Translated audio can be sent directly into voice chat through the HaloVoice Virtual Microphone, allowing other players to hear the translated voice in real time. HaloVoice can be used as a gaming voice translator across multiplayer games and gaming communities, including use cases such as a Valorant voice translator and Minecraft voice translator. For players using Discord, HaloVoice provides real-time voice translation for Discord. Simply select the HaloVoice Virtual Microphone as your microphone input in Discord to send your translated voice directly into the voice chat. Real-Time Voice Translator for Streaming HaloVoice provides real-time voice translation for streaming, allowing streamers to speak in one language while reaching audiences who speak another. For creators using OBS, HaloVoice provides real-time voice translation for OBS through the HaloVoice Virtual Microphone. Streamers can select HaloVoice as their microphone input in OBS and send translated speech directly into their stream. HaloVoice can also support multilingual live streaming workflows on platforms such as Twitch, helping creators communicate with audiences across languages without interrupting the broadcast. Real-Time Voice Translation for Online Meetings HaloVoice also provides real-time voice translation for Zoom, Slack, Microsoft Teams, and other communication platforms. Users can select the HaloVoice Virtual Microphone as their microphone input so other participants hear the translated voice directly during conversations. This makes HaloVoice useful for multilingual online meetings, remote teams, international collaboration, and everyday cross-language voice communication. HaloVoice supports practical language-pair use cases including Spanish to English voice translation, English to Korean voice translation, and other multilingual conversations. Key Features • Real-time voice translation • AI voice translator for multilingual conversations • Real-time voice translator for gaming • Real-time voice translation for streaming • Real-time voice translation for Discord and voice chat • Real-time voice translation for OBS and live streaming • Real-time voice translation for Zoom, Slack, and Microsoft Teams • AI voice cloning for translated speech • HaloVoice Virtual Microphone • Bilingual subtitles • Spanish to English voice translation • English to Korean voice translation • Support for Windows and macOS HaloVoice is built for gamers, streamers, creators, remote teams, and anyone who wants to communicate across languages while keeping conversations fast, natural, and voice-first.
BabelPhone is an AI app for real-time phone call translation, transcription, and recording. BabelPhone is an advanced AI app that translates phone calls in real time. It transcribes and records conversations with natural-sounding voice translation. The app uses VoIP calls, allowing users to dial any number locally or internationally without incurring extra charges on their mobile bill. BabelPhone also provides live transcriptions during calls and allows users to export a video recording complete with transcription.
An all-in-one AI-powered screen recorder and video editor designed exclusively for macOS. Dina is a professional macOS application that streamlines the entire video creation workflow into a single tool. It allows users to record their screen, webcam, and system audio, while providing advanced post-production features such as automatic cinematic zoom, smooth cursor movement, and transcript-driven editing. Designed specifically for the AI era, it leverages on-device Core ML models to generate captions and voiceovers locally, ensuring privacy. It features a native macOS experience built with SwiftUI and Metal, supporting high-quality exports up to 8K resolution without the need for recurring subscriptions.
All-in-one screen recorder with AI editing for demos, tutorials, and courses. Tella is an all-in-one online screen recorder for Mac & Windows, designed to help users create incredible product demos, tutorials, and courses. It simplifies video creation by allowing recording in small, manageable clips with speaker notes. Tella stands out with its AI video editing features, which enable users to remove filler words, silences, and edit videos like a document using text. It also offers transitions, zoom effects, customizable backgrounds, and powerful publishing options including instant sharing, 4K export, and embedding. Tella aims to provide a professional video creation experience without requiring advanced editing skills, positioning itself as a faster and more intuitive alternative to traditional video editing software and competitors like Loom.
AI-powered Chrome extension for transcribing and summarizing online meetings and media. Meetmemos is a Chrome extension that uses AI to transcribe and summarize online meetings, lectures, and media experiences from platforms like YouTube and Google Meet. Powered by OpenAI, it offers real-time, accurate transcriptions and smart summaries, turning lengthy content into digestible insights. It is designed to enhance productivity and information retention from online engagements.
AI-powered memory support, training & productivity app for cognitive differences. Recallify is your AI-powered memory companion, helping you to capture, recall, and enhance memories with cutting-edge AI. Seamlessly record text, audio, and video, and transform them into learning experiences while practising and improving recall. Recallify helps people with cognitive differences manage their day-to-day lives. Upload or record audio, video, text & PDF,  get ultra-accurate AI transcription and summaries,  train your memory and stay organised.
AI-powered tool to remove profanity from audio and video files. CurseCut is an AI-powered tool designed to detect and mute profanity and other unwanted keywords from audio and video files. It allows users to easily create family-friendly, professional content by filtering out offensive language. The tool offers customizable filtering capabilities, enabling users to define which words are considered inappropriate based on their specific needs and preferences.
Social Tooling: TikTok content analysis and transcription tool. Social Tooling is an innovative tool designed for creators, marketers, and casual browsers to navigate TikTok content. It uses advanced batch transcribe and AI analyse videos to help users explore and analyze competition, discover content ideas, and enhance their social media strategy.
Open source LMS for creating, managing, and delivering courses with AI support. ClassroomIO is a platform for bootcamps, individual educators, and training businesses that brings teaching and learning into one place while at the same time helping them be 10x more productive. It is an open source learning management system (LMS) for companies, offering a flexible, user-friendly platform for creating, managing, and delivering courses. Features include customizable LMS, AI support, simplified course management, and collaborative student and teacher community.
Perso Dubbing is an AI video dubbing platform that translates, dubs, and lip-syncs videos into 99+ languages. AI voice cloning preserves each speaker's tone and emotion, and multi-speaker detection handles up to 10 speakers per video. It reduces localization costs by up to 98% compared to traditional dubbing studios. Developed by ESTsoft and trusted by 450,000+ users. Perso Dubbing is an AI-powered video dubbing and translation platform that localizes content into 99+ languages in minutes, with speech recognition in 100+ languages. Teams upload a video, select target languages, and receive a studio-quality dubbed version — complete with lip-sync and voice cloning that preserves the original speaker's tone, accent, and emotion. Key capabilities: • AI Voice Cloning — Matches the original speaker's voice, accent, and emotional tone across all dubbed tracks • AI Lip Sync — Aligns translated audio with on-screen mouth movements for natural viewing • Speech-to-Text — Speech recognition in 100+ languages • Audio Separation — Splits voice and background tracks • Auto Subtitle Generation — Creates and exports subtitles automatically • Real-Time Script Editor — Review and refine translations before final export • Multi-Speaker Support — Detects and dubs up to 10 speakers in a single workflow Built for marketing teams, e-learning creators, enterprise L&D departments, and media publishers expanding into global markets. Enterprise plans include API access, advanced security controls, and dedicated support. Developed by ESTsoft (est. 1993, KOSDAQ: 047560) — ISO/IEC 27001 and KISA ISMS certified.
DesiVocal is a free AI voice generator for HD voice overs in multiple languages. DesiVocal is a free text-to-speech and AI voice generator that creates HD AI voice overs in multiple languages. It caters to youtubers, publishers, and media houses, offering premium AI voice overs in seconds. It also provides a speech-to-text feature.
Virtual studio for high-quality remote podcast and video recording and editing. Riverside is an all-in-one podcast and video studio for businesses and content creators who demand professional quality without technical complexity. Podcasters, video creators, and businesses across the globe are producing studio-quality content in a fraction of the time it would take with traditional methods.
Transcribe audio, video, YouTube links, and recordings into accurate, searchable text. Audio Converter AI makes it easy to transcribe audio to text online without installing software. Convert interviews, meetings, lectures, podcasts, videos, voice recordings, and YouTube content into accurate, searchable, and editable transcripts. The AI transcription engine supports more than 200 languages and can automatically detect the source language. Speaker recognition helps separate conversations, while timestamps make it easy to find important moments in long recordings. After transcription, you can use AI-generated summaries and smart notes to understand content faster and create reusable materials. Audio Converter AI supports popular formats including MP3, MP4, M4A, WAV, WEBM, MOV, AIFF, OPUS, FLAC, AVI, MKV, FLV, and 3GPP. Files can be up to 10GB, with multiple tasks supported in the queue. Your content is processed in encrypted environments, and files are automatically removed after processing.
API-based voice and SMS platform for communication, data capture, AI analysis, and CRM integration. Callr is an API-based voice and SMS platform that seamlessly integrates into your product, enabling powerful communication, data capture, AI analysis, and CRM integration. It allows turning conversations into actionable data.
Conversation Experience Platform with Generative AI and Speech Recognition. Seasalt.ai provides a Conversation Experience Platform with Generative AI and Speech Recognition. It aims to help businesses capture, generate, and understand all text and voice conversations with customers, enabling natural, personalized, and actionable interactions.
VetRec is an AI scribe for veterinarians that autogenerates medical records quickly. VetRec is an AI Scribe exclusively for Veterinarians. It autogenerates Medical Records in 30 seconds after a consultation is done. Veterinarians can focus on what matters the most, the pet! VetRec is HIPAA compliant. It is an AI assistant for veterinarians to work more efficiently before, during, and after their consults.
AudioShake uses AI to split audio recordings into stems for various interactive and customizable uses. AudioShake's AI can split any recording – from music to film to UGC content – into its stems, making audio more interactive, customizable, and accessible. It offers services like dialogue, music & effects separation, lyric transcription & alignment, and instrument stem separation. Use cases include mixing & mastering, localization & captioning, interactive sync licensing, audio analysis, A/V editing, fan engagement, and copyright compliance.