
WhisperTranscribe is an AI-powered transcription tool that converts audio and video into text with 95% accuracy, supporting 55+ languages. It also offers features like speaker detection, chat with audio, and content creation from transcripts.
Category
All 114 AI Speech Recognition tools in the index, newest first. Each entry says what the tool does, which platform it runs on, and links straight to it.
Category
114 AI Speech Recognition tools indexed · page 2 of 4

WhisperTranscribe is an AI-powered transcription tool that converts audio and video into text with 95% accuracy, supporting 55+ languages. It also offers features like speaker detection, chat with audio, and content creation from transcripts.

Transcript LOL is an AI-powered transcription tool that converts audio and video files into text with high accuracy. It is used by professionals, content creators, and organizations for transcribing interviews, podcasts, meetings, and more.

Listen Monster converts audio and video files to text with high accuracy. It supports many languages and offers unlimited transcription plans for heavy users.

EquiLoomPRO is an AI-powered investment platform that uses an algorithmic bot to automate trading strategies. It is designed for individuals looking to enhance their investment techniques and potentially increase their earning potential.

Pitch Patterns is a tool that helps users analyze and improve their sales pitches. It is designed for sales professionals and teams looking to refine their communication strategies.

MemFlow is a conversational research assistant for scientific discovery. It learns from every interaction to build a dynamic knowledge base, automating tedious tasks and identifying knowledge gaps.

OneAudio converts audio recordings into clean, well-structured notes. It is designed for individuals who want to capture and organize ideas by speaking or uploading audio files.

Momentary is an AI-powered voice journaling app that turns spoken thoughts into written journal entries. It is designed for individuals seeking to track their emotions and engage in self-reflection with the help of an AI mentor.

VOMO AI is a meeting notes and audio transcription tool that converts recordings, uploads, and YouTube links into accurate transcripts with AI summaries and action items. It is designed for teams, creators, and students who need to capture and organize meeting content efficiently.

VOMO is an iOS app that records, transcribes, and summarizes meetings, lectures, and other audio. It is designed for professionals who need automated meeting notes and AI-powered insights.

Speechless is an iOS app that transcribes and translates audio using OpenAI's Whisper API. It allows users to import audio from the app or share sheet and supports integration with apps like WhatsApp, Messages, and Voice Memos.

XSTAR168 is an online slot platform that aggregates top slot game providers. It is designed for slot enthusiasts who want a smooth gaming experience and easy transactions.

JimakuAI transforms long-form videos into Japanese subtitles, short clips, and carousel posts. It serves enterprise teams, content creators, and media brands.

TalkNotes is an AI-powered voice note app that transcribes and structures voice recordings into text. It is designed for professionals, content creators, and students who want to save time on note-taking.

SpeechPulse is a speech recognition and translation tool for Windows and macOS. It allows users to dictate text into any application, supports offline transcription for privacy, and offers features like speaker diarization and subtitle generation.

Speech Meter is a web-based tool that analyzes your accent and provides instant feedback on your English pronunciation. It is designed for individuals looking to improve their English speaking skills.

TakeNote is an AI-powered meeting notes and document extraction tool for UK regulated financial advisers. It generates FCA-compliant suitability records, auto-fills provider forms, and provides team-scoped access with a full audit trail.

Freed is an AI medical scribe that turns patient conversations into accurate clinical notes and offers coding, decision support, and front desk assistance. It is designed for independent clinics and small practices.

ChatDox AI 2.0 is an AI-powered tool that lets users ask questions and get answers from documents, YouTube videos, websites, audio, and video files. It is designed for students, researchers, teachers, businesses, and professionals.

MAIA is a Chrome extension that provides a personal AI assistant for summarizing, generating, explaining, simplifying, translating, and transcribing content. It is designed to be usable, accessible, and affordable, with a pay-as-you-go pricing model.

Voice to Text is a tool that converts spoken words into written text. It is designed for users who need to transcribe audio quickly and easily.

Autocalls is an all-in-one platform for deploying AI voice agents that make and receive phone calls autonomously. It is designed for businesses to automate appointment booking, customer support, and cold calling.

Vaanee AI Engine is a generative AI voice platform that offers text-to-speech, voice cloning, speech-to-speech translation, and AI video dubbing. It is designed for content creators, filmmakers, and media professionals who need realistic, multilingual voiceovers.

TranscriptMate is an AI-powered transcription service that converts audio and video files to text with 98% accuracy. It offers speaker identification, timestamps, and multi-language support, and is designed for professionals such as journalists, researchers, and content creators.

Relevant is a podcast production tool that provides real-time content suggestions and topic detection. It is designed for podcasters who want to enhance their conversations with relevant web content and streamline their workflow.

VideoToTextAI is a free AI transcript generator that converts Instagram Reels, TikTok videos, meetings, podcasts, and uploaded audio or video files into text. It is designed for creators and teams who need to repurpose spoken content into social posts, show notes, and other written formats.

Robo Translator is a machine translation service built on OpenAI and Azure Cognitive Services. It helps users localize content such as audio, video, text documents, and software files into multiple languages.

Insight Video IA is an AI-powered tool that converts video lessons into educational resources such as e-books, summaries, quizzes, and mind maps. It is designed for teachers and educators to enhance their teaching materials and student engagement.

Felo Subtitles is a real-time multilingual subtitle and transcription tool that works with meeting platforms like Zoom, Teams, and Google Meet. It provides instant translation, meeting summaries, and customizable dictionaries for professionals.

Malloy Studio is an AI-powered motion graphics generator that turns text prompts into customizable templates for videos. It is designed for video editors, social media creators, and marketing teams who want to add professional-looking animations without learning After Effects.

Aispect turns live audio into visual images in real time. It is designed for events, webinars, meetings, and other live audio sources.
Real-time AI copilot for interviewees

VoiceRec is an AI-powered voice recorder app for iOS and iPadOS that records audio and transcribes it to text. It is designed for users who need to capture and organize voice recordings, such as meetings, lectures, or presentations.

Gladia is an AI audio infrastructure that provides transcription and audio intelligence through a single API. It is designed for developers building voice products who need accurate, multilingual transcription with features like speaker detection and entity recognition.

Ello is an AI-powered reading and math app for children aged 4–9. It adapts to each child's needs in real time, offering personalized lessons and books.

Botjet is a conversational AI platform for businesses to build and deploy chatbots across web, IoT, and mobile. It offers technologies like conversation engine, deep learning, speech recognition, and speech synthesis.