AI Speech-to-Text 350

AI-powered transcription service for audio and video to text conversion. VideoToWords AI is an AI-powered transcription service that converts audio and video files into accurate written text. It supports over 98 languages and various file formats, offering features like automatic transcription, editing, and exporting to formats like TXT, DOCX, and SRT. It caters to a wide range of users, including journalists, students, researchers, and content creators, providing a fast and efficient way to transcribe audio and video content.
AI audio and video processing platform with tools for transcription, translation, and editing. RecCloud is a leading AI audio and video processing platform that offers a range of tools for content creation and editing. It includes features like AI speech-to-text, AI subtitles, AI text-to-speech, and AI video translation. The platform is designed to be user-friendly and accessible online.
Free online tool to transcribe audio and video to text with translation. FreeSubtitles.AI is a free online tool that transcribes audio and video to text, offering automatic translation. It allows users to upload files or use a media downloader for content from various websites. The platform provides both free and paid options, with increased limits and features for paid users, such as larger file sizes, longer durations, and more accurate transcription models.
AI-powered subtitle translation and audio transcription web application. GPT Subtitler is a web app for fast, accurate subtitle translations using LLM like OpenAI, Claude, or Gemini. It also offers audio transcription using Whisper AI. Users can translate subtitles between multiple languages and transcribe audio, transforming their workflow with efficiency and precision.
AI-powered voice journal for mental health and self-growth with AI mentorship. Momentary is an AI-powered journal that helps users preserve their treasured moments with voice. It allows users to speak their thoughts, and AI crafts voice journals, capturing titles, themes, and moods. Users receive personalized AI mentorship for emotional and cognitive growth.
AI tool for turning videos into transcripts and summaries SummarizeVideoToText is an online AI tool that converts YouTube, TikTok, and Instagram videos into searchable transcripts, concise summaries, timestamped chapters, mind maps, and interactive Q&A. Users can paste a video link and receive text-based insights without downloading software or extensions.
AI Teacha empowers educators with AI tools for efficient teaching and personalized learning. AI Teacha: Empowering educators with AI tools for efficient lesson planning, assessments, curriculum design, and content creation. Revolutionize education, personalize learning, and engage students effectively. AI Teacha streamlines lesson planning, assessment creation, and curriculum development, saving teachers valuable time that can be dedicated to instruction and student support.
Voice-based story builder for kids to spark imagination and foster creativity. Stories With Dory is a voice-based story builder that sparks your child’s imagination, adapts to their interests, and fosters creativity with captivating tales, taking kids into a world of imagination and reading. It's an interactive story builder that transforms creativity into magical adventures, making storytelling fun and engaging. The app helps children become master storytellers by turning screen time into story time to boost creativity, improve vocabulary, and enhance storytelling skills in a fun and engaging way.
All-in-one CRM for online sales and call centers. SalesRender is an all-in-one CRM designed for online sales and call centers. It provides tools to automate dialers, track team productivity, resell leads to CPA networks, and audit calls with AI, giving users full control over their sales pipeline without requiring extra development work. The platform caters to online stores, call centers, and CPA networks, offering expert support throughout implementation. It's scalable for enterprises managing millions of orders and hundreds of call center operators, while also being beneficial for beginners. Key functionalities include call management, telephony integrations, process automation, operator performance monitoring, analytics tools, AI-powered call auditing, flexible automation with triggers, comprehensive order management, a CPA module, multi-channel chat integrations, warehouse and item management, and robust security features with customizable roles and detailed logging.
MacOS app for email management and task automation with AI-powered workflows. Inbox AI is a MacOS application designed to manage email and automate everyday tasks using custom AI-powered workflows. It offers both cloud-based and privacy-first on-device AI options. It natively connects with Apple Reminders and integrates with other apps through API or file-based commands. Users can have conversations with AI, automate tasks, process emails, and capture information via voice. It allows users to build custom AI-powered assistants that can answer questions and perform actions, such as drafting emails, remembering information, discussing text, rewriting content, replying to messages, searching the web, opening apps, and transcribing text.
Open-source VoIP phone with AI for call transcription, summarization, and CRM integration. 008 is an open-source VoIP phone with advanced AI capabilities. It transcribes conversations, summarizes them, and extracts business KPIs. It seamlessly logs events and interactions, and transfers call data to preferred CRMs and tools. It aims to revolutionize customer service by automating calls and providing valuable insights.
On-device voice assistant for conversational AI coding SKI is an on-device voice assistant for AI coding agents such as Claude Code, Cursor, Codex, Gemini CLI, Windsurf, and OpenClaw. It lets developers speak naturally to their coding agent, have the agent build and manage code, and hear its responses aloud. Unlike dictation tools, SKI provides a full two-way voice conversation with local speech-to-text, neural voice synthesis, full-duplex interruption, agent status updates, multi-project support, and optional meeting participation. Voice data, transcripts, and local meeting recordings remain on the user's computer and are not uploaded.
An AI voice assistant and dictation tool for high-speed cross-app productivity on desktop. NovaVoice is an AI-powered Voice OS for desktop environments designed to accelerate productivity by replacing typing with high-speed voice commands. It allows users to dictate text at over 200 words per minute, which is roughly four times faster than average typing speeds. Beyond simple speech-to-text, NovaVoice functions as a cross-app copilot that can reformat text into various styles, execute real actions like sending messages or emails via voice, and provide instant AI-assisted answers based on screen content. It integrates a 'Terms Dictionary' for frequently used information and supports macOS, Windows, and Linux.
AI-powered content moderation platform for protecting brands and user experience. Lasso Content Moderation is an AI-powered solution designed to streamline the content moderation process. It offers an intuitive dashboard and a seamless API integration, providing a full moderation solution right out of the box. Lasso helps protect brands and safeguard user experiences by removing harmful content in a scalable and cost-effective way. It supports text, image, video, and audio moderation, and integrates with popular platforms like Slack, TalkJS, Stream, and more.
Desktop AI assistant for real-time interview answers and practice analytics SubcueAI is a desktop AI interview assistant for macOS and Windows that provides real-time answer suggestions, live transcription, dual audio capture, and post-session interview analytics.
Undetectable AI meeting and interview assistant with local or cloud-based processing. Natively is an advanced hybrid AI meeting and interview assistant designed to provide real-time support during video calls. It functions as an undetectable overlay that can run 100% locally via Ollama for maximum privacy or connect to cloud-based LLMs like OpenAI, Anthropic, and Gemini using your own API keys. The tool captures live audio to generate perfect meeting notes, provides real-time answers to interview questions, and offers specialized modes for technical interviews, sales, and recruiting. It integrates seamlessly with platforms like Zoom, Google Meet, and Microsoft Teams without the use of visible meeting bots, making it completely invisible to other participants during screen shares.
HIPAA-compliant AI agent for healthcare, using OpenAI with data security. CompliantChatGPT is an AI Agent designed to assist with healthcare-related tasks while ensuring patient data remains safe, secure, and HIPAA Compliant. It leverages OpenAI's GPT models with added security measures like PHI tokenization to maintain compliance.
Sully.ai offers AI agents for healthcare to automate tasks and save doctors time. Sully.ai provides AI Agents for healthcare, aiming to save doctors time by offering an AI team including a Nurse, Receptionist, Scribe, Med Asst, Coder & Pharmacy Tech. These agents seamlessly handle tasks from check-in to prescriptions.
Local AI meeting assistant for private transcription and summaries Meetily AI is a privacy-first, open-source meeting assistant that records, transcribes, and summarizes meetings entirely on your device. It works bot-free with Zoom, Microsoft Teams, Google Meet, and other platforms, while keeping audio and data local under your control.
Versive is an AI-powered user research platform for faster insights and decision-making. Versive is an all-in-one user research platform that helps companies conduct and analyze research faster, using AI. It assists in designing, moderating, and synthesizing research, leveraging the power of AI to accelerate the process from questions to insights. Versive offers tools for collecting richer insights through AI-moderated studies, flexible surveys, AI-moderated interviews, usability tests, and translation services. It also accelerates analysis by turning transcripts into shareable insights with AI assistance.
AI-powered platform for audio and video transcription and translation. Cockatoo is an AI-powered platform that transcribes audio and video files into text and subtitles. It offers high accuracy, supports over 90 languages, and provides unlimited transcripts. The platform is designed to be simple and easy to use, allowing users to convert audio and video to text in seconds and export to popular formats like docx, pdf, and srt. Cockatoo emphasizes privacy and security, ensuring user data is protected with state-of-the-art cryptography and never shared with third parties.
AI-powered voice-to-text tool for fast, structured, and professional dictation. VoiceDash is an AI-powered voice typing tool designed to convert speech into structured, professional text instantly. It integrates with existing applications on Mac, Windows, and mobile devices to boost productivity by eliminating filler words and correcting grammar in real-time. The tool is designed to work seamlessly across various platforms, allowing users to communicate at the speed of thought for tasks such as client notes, reports, emails, and manuscript drafting.
AI-powered voice cloning, text-to-speech, and speech-to-text platform. Voicv is a cutting-edge voice cloning platform that transforms your voice into a digital asset in minutes, supporting multiple languages and zero-shot learning. It offers advanced AI-powered voice cloning, text-to-speech (TTS), and speech-to-text (ASR) services. Users can create, transform, and convert audio with cutting-edge technology, supporting multiple languages and emotions.
AI voice dictation app 4x faster than typing, converting messy speech into clear text. Genspark Speakly is an AI voice dictation application designed to convert spoken language into clear, polished messages, emails, and writings. It is marketed as being 4x faster than typing. The app integrates advanced AI features like Auto-Edits (which remove filler words, fix typos, and format text) and Custom Instructions (allowing users to define how their voice should be transformed, such as translation, CLI commands, or professional rewrites). It works across more than 100 applications and supports over 100 languages, making it a versatile productivity tool.