Transcription 152

AI meeting note taker and screen recorder for increased productivity and async communication.
Bubbles is an AI meeting note taker and screen recorder designed to increase productivity. It offers automated note-taking, AI action items, and summaries for meetings. Users can also switch to video conversations and screen recording for asynchronous communication. Bubbles integrates with existing tools and automatically updates Google Calendar with meeting notes, turning every meeting into actionable steps and helping users skip unnecessary meetings.

Open-source VoIP phone with AI for call transcription, summarization, and CRM integration.
008 is an open-source VoIP phone with advanced AI capabilities. It transcribes conversations, summarizes them, and extracts business KPIs. It seamlessly logs events and interactions, and transfers call data to preferred CRMs and tools. It aims to revolutionize customer service by automating calls and providing valuable insights.

Automatic transcription and subtitle platform with AI-powered tools and multi-format support.
ScriptMe is an automatic transcription and subtitle platform developed by post-production experts. It's designed to work with Avid Media Composer and many other tools. ScriptMe transcribes audio/videos and adds subtitles quickly, supporting +31 languages. It allows users to transcribe, subtitle, translate, and export in many formats. The platform offers features like AI transcription, subtitle generation, and translation services for various use cases, including YouTube videos, podcasts, interviews, meetings, and academic research. ScriptMe also provides enterprise solutions for TV & Media transcription, movie subtitling, and TV show transcription.

Free online tool to transcribe audio and video to text with translation.
FreeSubtitles.AI is a free online tool that transcribes audio and video to text, offering automatic translation. It allows users to upload files or use a media downloader for content from various websites. The platform provides both free and paid options, with increased limits and features for paid users, such as larger file sizes, longer durations, and more accurate transcription models.

AI audio and video processing platform with tools for transcription, translation, and editing.
RecCloud is a leading AI audio and video processing platform that offers a range of tools for content creation and editing. It includes features like AI speech-to-text, AI subtitles, AI text-to-speech, and AI video translation. The platform is designed to be user-friendly and accessible online.

Browser-based AI subtitling, transcription, and translation platform supporting 90+ languages.
Banva is a browser-based subtitling product that allows users to generate automated subtitles for 50+ languages, edit generated subtitles, and style subtitles. It's an AI-powered media localization platform with support for more than 90 languages and all the common video, audio and subtitle formats. Banva fulfills all subtitling needs right in the browser with fast and accurate automatic subtitle generation and a complete subtitle editor suite.
Audio & video transcription tool with AI meeting summary and action items generation.
Mictoo is an audio and video transcription tool that converts audio to text automatically. It allows users to record audio or upload files to get real-time transcription. Mictoo also uses GPT Open AI to generate meeting summaries, action items, and follow-ups that can be shared with colleagues. It helps users take meeting notes easier, freeing up their minds to engage positively in meetings and enhance productivity.

AI-powered voice collaboration platform for transcription, summarization, and actionable insights.
Vocol is an all-in-one voice collaboration platform powered by AI, designed to boost work efficiency by turning voice and data into actionable insights. It transcribes and summarizes meetings, supports multilingual transcription (Chinese, Japanese, and English), and integrates with tools like Teams.
Free web-based tool to transcribe and summarize MP3 files using AI.
WebWhisper is a FREE web-based alternative for MacWhisper that allows you to transcribe and summarize MP3 files effortlessly. It utilizes advanced AI models like GPT-3.5, GPT-4, and Claude to get accurate transcriptions and concise summaries.

AI-powered meeting note-taking and transcription tool.
Minutes AI automates meeting audio notes by instantly creating formatted notes and transcriptions from live audio, uploaded audio files, or imported YouTube links. Users can chat with their audio to extract key insights, list action items, and more. It's designed to be reliable, simple, private, and powerful, helping users never take notes manually again.

AI-powered voice notes app for creating content from audio in 90+ languages.
NoteGen is an AI-powered voice notes app that instantly turns your ideas into effective content. It supports 90+ languages and allows you to effortlessly create journals, notes, scripts, posts, call summaries, and more in one click. You can record or upload audio for note-taking, call summarizing, journaling, creating posts, content scripts, and more.
Unifies speech recognition across 1,600+ languages using AI and LLM-enhanced decoders.
Omnilingual ASR is an advanced automatic speech recognition technology that unifies speech recognition across a vast number of languages, scaling from dozens to over 1,600 natively and extending to 5,000+ via few-shot prompts. It achieves this by combining wav2vec-style self-supervision, LLM-enhanced decoders, and balanced multilingual corpora to learn language-agnostic acoustic patterns. This website serves as a comprehensive knowledge base, detailing its research breakthroughs, current technologies, datasets, implementation strategies, and deployment guidance for achieving omnilingual reach in a single model.

Multilingual Speech-to-Text API with high accuracy in 14 languages.
SpeechFlow is a multilingual Speech-to-Text API that offers state-of-the-art accuracy in 14 languages. It converts sound to text, speech to text, and audio to text with high accuracy. SpeechFlow supports both cloud and on-prem deployment.

macOS app converting speech to text with ChatGPT, speeding up writing.
WhisperWizard is a macOS application that transforms spoken words into written text with the help of ChatGPT. It speeds up writing workflows by allowing users to speak instead of type, capturing ideas instantly and accessing old recordings. It also offers custom ChatGPT prompts to edit recordings and create templates for routine tasks.

Privacy-focused macOS app for instant on-device dictation and AI-powered text processing.
Spoke is a native macOS dictation application that provides instant voice-to-text transcription directly into any text field. It runs a high-performance local speech recognition model via CoreML, ensuring that audio never leaves the device for maximum privacy. Users can hold a customizable keyboard shortcut to speak and see their words appear instantly at the cursor. The app also features 'AI Skills,' allowing users to process transcriptions on the fly for tasks like translation, grammar correction, or summarization using their own API keys from providers like OpenAI, Anthropic, or local models via Ollama.

AI-powered transcription service converting audio and video to text with high accuracy and speed.
Yescribe.ai is an AI-powered transcription service that precisely converts audio and video files to text. It supports multiple formats and 98 languages, providing fast, accurate, and secure transcriptions. Users can simply upload their file and enjoy the convenience of AI-powered transcription of audio/video into text, helping them focus on what's really important. Yescribe.ai offers features like 99.9% accuracy, global language coverage, extended support for up to 5-hour uploads, rapid transcription with instant results, AI summaries for insightful overviews, and private & secure data handling.

WhisperUI: Affordable text-to-speech and speech-to-text service using OpenAI Whisper API.
WhisperUI is a text to speech and speech to text service powered by OpenAI Whisper API. With WhisperUI you can use your OpenAI api keys to get affordable text to speech and speech to text services. It allows users to convert audio files to text and SRT files using OpenAI Whisper Speech to Text.

All-in-one AI content generator for text, images, code, and voice solutions.
Maximus AI is an all-in-one AI content generator that combines the power of an AI writer, AI chatbot, AI code generator, speech-to-text, and AI image generator into a single platform. It aims to revolutionize content creation by providing AI-driven text, image, code, and voice solutions, offering a cost-effective way to generate content quickly and increase conversion rates.

AI-powered platform for X Spaces transcription, analysis, and Q&A.
XSPACESTREAM is an AI-powered platform that transforms X Spaces (Twitter Spaces) into actionable insights. It provides complete transcripts, summaries, and AI Chat capabilities to answer questions about the content. The platform offers real-time transcription, AI-powered analysis, and interactive Q&A, enabling users to discover what's being said in X/Twitter Spaces.

AI tool for podcasters and content creators to automate content repurposing and creation.
Write Panda is an AI-based tool designed for podcasters and content creators to automate time-consuming tasks such as creating shownotes, timestamps, titles, mentions, blogs, newsletters, and tweets. It repurposes content into various formats, including blogs, newsletters, tweet threads, and auto-captioned viral clips, helping users reach a wider audience and grow their subscribers, listeners, and followers without extensive manual content creation.
AI-powered online video editing platform with auto subtitling, translation, and resizing.
Nova A.I. is an online video editing platform powered by AI, offering features like auto subtitling, translation, transcription, and video resizing for various digital platforms. It includes tools for adding transitions, CTAs, and more. It also provides a computer vision video search engine. Nova A.I. aims to be simple to use, save time, and require no installation.

Video translation and dubbing service with voice cloning.
Translate This Video is a service that converts English-speaking videos into over a dozen languages, allowing users to share their content with a global audience. The service uses voice cloning technology to dub videos in new languages with voices that sound like the original speakers. It also provides instant transcripts in multiple languages and offers transcript editing capabilities.

AI-powered platform for dubbing videos in multiple languages with lip-sync.
Speax is an AI-powered platform designed for dubbing videos in multiple languages. It offers a user-friendly interface, batch processing capabilities, and fast results, enabling users to easily reach a global audience. The platform supports uploading, managing, and downloading videos. Speax provides AI-powered video dubbing and voiceovers with lip-sync, multilingual accuracy, and natural AI voices, optimized for global audiences.

AI-powered medical scribing and documentation tool for healthcare professionals.
Mediscribe Pro is an AI-powered medical scribing, charting, and documentation tool designed for healthcare professionals. It aims to reduce paperwork and burnout by generating dictations, transcriptions, and chart notes. The tool is HIPAA & PIPEDA compliant and offers features like medical templates, code extraction, and EMR integration.