transcription 9

AI-powered offline voice-to-text app for macOS, supporting 100+ languages. superwhisper is an AI-powered voice-to-text application for macOS that allows users to dictate emails, send messages, and take notes at speeds up to three times faster than typing. It operates completely offline, ensuring privacy and security as data never leaves the user's device. superwhisper supports over 100 languages and offers features like literal punctuation control in its Pro version.
AI-powered platform for studio-quality video and podcast creation, editing, and distribution. Podcastle is the easiest way to create studio-quality videos and podcasts. Record, edit and distribute content directly in your browser using AI-powered tools. It's a one-stop shop for broadcast storytelling, great for podcasters or anyone who deals with long-form video creation. Studio-quality recording, AI-powered editing, and seamless exporting – all in a single web-based platform.
AI-powered memory support, training & productivity app for cognitive differences. Recallify is your AI-powered memory companion, helping you to capture, recall, and enhance memories with cutting-edge AI. Seamlessly record text, audio, and video, and transform them into learning experiences while practising and improving recall. Recallify helps people with cognitive differences manage their day-to-day lives. Upload or record audio, video, text & PDF,  get ultra-accurate AI transcription and summaries,  train your memory and stay organised.
Intelligent audio & video to text converter NeatScribe harnesses advanced AI and speech recognition capabilities to provide rapid, high-quality transcription of audio and video content. The platform specializes in transcribing podcasts, meeting recordings, interviews, lectures, and other media formats into editable text. Users simply upload their files and receive precise transcripts in approximately 60 seconds. With support for exporting to TXT, PDF, DOCX, SRT, VTT, and other popular formats, it serves content creators, journalists, academics, students, and professional teams looking to streamline their media-to-text workflows.
deAPI is an OpenAI-compatible inference API for open-source models — image, video, audio, music, embeddings, OCR. $5 free, no card. Generate images, create videos, clone voices, transcribe YouTube — through one AI inference API and one billing account. deAPI serves 25 open-source AI models across every non-LLM modality, so you don't need to juggle five providers for a single product. ─── TWO WAYS TO CONNECT ─── • OpenAI-compatible endpoint — drop-in for the OpenAI Python and Node SDKs. Change base_url and api_key, keep everything else. • Native REST v2 API — covers everything the OpenAI spec doesn't: video, music, OCR, voice cloning, voice design, and more. ─── IMAGE · 7 MODELS ─── Flux.1 Schnell, FLUX.2 Klein, Z-Image-Turbo, Z-Anime Distill, image editing, background removal, AI upscaling. From $0.00136/image. ─── VIDEO · 7 MODELS ─── Text-to-video, image-to-video, audio-to-video with LTX. Character replacement with Wan2.2-Animate. Video upscaling up to 4x. From $0.00174/clip. ─── AUDIO · 9 MODELS ─── • TTS — Kokoro (40+ voices, 7 languages), Chatterbox (23 languages), 3× Qwen3 TTS. Voice cloning from a 5-15s clip. Voice design from text. From $0.77/1M chars. • Music — ACE-Step: full tracks with lyrics, BPM, key, time signature, style transfer. • Transcription — WhisperLargeV3, 80 MB uploads (3x OpenAI), URL ingest from YouTube, TikTok, Twitch, Kick, X. From $0.021/hour. ─── TEXT · 2 MODELS ─── OCR with Nanonets. Embeddings with BGE-M3 (1024 dims, 8192 tokens) for RAG and semantic search. From $0.000068/1K tokens. ─── PRICING ─── Pay-as-you-go, no subscription. $5 free credit on signup, no card required. Any top-up unlocks Premium: 300 RPM per endpoint, no daily cap. ─── FOR DEVELOPERS ─── Python SDK, n8n node, MCP server for Claude/Cursor/ChatGPT. Webhooks (HMAC-signed), WebSocket live previews, or polling. Distributed verified worker network — data in RAM only, no disk writes.
AI dictation that turns natural speech into clear, formatted text in any app. Aqua Voice is system-wide AI dictation for macOS, Windows, and iPhone. It converts natural speech into clear, polished text in any app, with strong technical-vocabulary accuracy, custom instructions, and a synced personal dictionary. Aqua supports 49 languages and is designed for fast writing, productivity, transcription, and speech-to-text workflows. It offers transparent freemium pricing and public privacy and security documentation. Aqua Voice, Inc. is SOC 2 Type II compliant.
AI meeting assistant and notetaker with automated CRM and no meeting bot. Sonnet AI is an end-to-end AI meeting assistant and notetaker that offers no-bot audio recording, automatic join meeting notifications, transcription, custom notes, action items, and auto-updating CRM—all without a meeting bot.
Accurate and affordable human-verified transcription services with AI enhancement. Scribie provides accurate and affordable human-verified transcription services. They offer transcription, audio to text conversion, and video transcription with 99% accuracy. Scribie uses a human-in-the-loop transcription and formatting service, combining AI tools with human expertise. They cater to various industries, including legal, academic, video, sermon, podcast, marketing, and audio.
Audio to text conversion service powered by OpenAI, supporting multiple languages and formats. Audio2Text is a service that converts audio to text with high accuracy, supporting multiple languages and audio file formats. Powered by OpenAI's Whisper AI, it offers both free and paid options, with the paid versions providing higher transcription quality and faster processing times. Users can transcribe audio files and export them in various formats like TXT, PDF, and SRT, making it suitable for creating subtitles and other text-based content.