getapp-logo

App comparison

Add up to 4 apps below to see how they compare. You can also use the "Compare" buttons while browsing.

GetApp offers objective, independent research and verified user reviews. We may earn a referral fee when you visit a vendor through our links. 

Top Rated Speech Recognition Software with Open source - Page 4

Last updated: August 2026

1 filter applied

Features


Integrated with

No filters available


Pricing model


Devices supported


Organization types


User rating


115 software options

TENIOS Voice API logo

Integration platform for integrating telephony applications

learn more
TENIOS Voice API facilitates the seamless integration of speech services into your cloud telephony using standard web technologies. This API includes a variety of functions that enable software applications to initiate and receive calls, eliminating the need for developers to handle TK technologies.

Read more about TENIOS Voice API

Users also considered
BeLora Connect logo

Voice translation for calls and meetings

learn more
Connect is an AI-powered voice translation app offering real-time speech interpretation in over forty languages for Zoom, Google Meet, Teams, and Slack. It uses voice cloning to preserve tone and emotion, operates with ~500ms latency, and includes features like speaker identification, context-aware accuracy, and end-to-end encryption with local transcript storage.

Read more about BeLora Connect

Users also considered
VoFact logo

AI voice invoicing for self-employed workers

learn more
Our customized speech recognition operates via WhatsApp voice notes. The AI accurately transcribes spoken job details, calculates totals, and instantly formats 2026-compliant electronic invoices. This hands-free approach allows freelancers to bill clients quickly without manual data entry

Read more about VoFact

Users also considered
DokuDachs logo

AI-powered therapy session documentation tool

learn more
DokuDachs is an AI documentation tool for psychotherapists to record, transcribe, and summarize sessions. It provides real-time transcription with speaker recognition, generating structured summaries linked to transcript locations. Data is encrypted, stored on GDPR-compliant European servers, and secured with zero-knowledge architecture. Audio files are not stored permanently.

Read more about DokuDachs

Users also considered
Irma logo

Cloud-based and AI-enabled meeting notes tool

learn more
Irma is a cloud-based AI meeting assistant that helps automatically capture meeting notes.

Read more about Irma

Users also considered
Writhere logo

On-device voice to text for macOS

learn more
Speak naturally in any macOS app and watch it become clean, punctuated text instantly. Writhere's on-device AI handles accents, filler words, and even lets you dictate in over 90 languages, or translate as you speak, all without sending audio anywhere.

Read more about Writhere

Users also considered
Wavel logo

Full Stack Voice AI Solutions for Videos And Localization

learn more
Wavel is an AI-powered video assistant that uses advanced text-to-speech and speech-to-text technology to create captions, subtitles, and dubbing in over 40+ languages. It offers voiceover customization with 250+ emotions and integrates with popular platforms like YouTube and Vimeo

Read more about Wavel

Users also considered
Dictalogic logo

Cloud speech recognition solution

learn more
With the use of digital transformation, we allow a voice to text conversion on the fly, where you just record audio and send it to transcribe as you normally would and the audio converts to text before it reaches the transcriber. We have multiple options on assignment for you to explore.

Read more about Dictalogic

Users also considered
Speech Recognition Cloud logo

Speech recognition for doctors, professionals & students

learn more
Speech Recognition Cloud is cloud-based speech recognition software for doctors, professionals and students. Fast, high-accuracy speech-to-text dictation in Windows apps and the browser. Free option available, plus specialised Medical for clinical terminology and workflows.

Read more about Speech Recognition Cloud

Users also considered
SpeechWrite 360 logo

Speech recognition and mobile dictation software

learn more
SpeechWrite 360 is a cloud-based dictation and voice recognition workflow solution designed to meet the needs of modern professionals requiring flexible and mobile working capabilities. Hosted in the secure Amazon Web Services cloud infrastructure, SpeechWrite 360 requires no onsite servers or IT resources. Users always have access to the latest software version with no additional upgrades or maintenance.

Read more about SpeechWrite 360

Users also considered
Ebby logo

Cloud-based transcriptions software

learn more
Ebby helps lawyers, podcasters, journalists, researchers, and academic professionals convert audio recordings into text documents using AI technology. The built-in editor automatically synchronizes and plays audio or video files with text data, letting users review and edit transcripts in real-time.

Read more about Ebby

Users also considered
silenis.online logo

AI video dubbing in multiple languages

learn more
Silenis is an AI-powered video dubbing platform that translates videos into 40+ languages while preserving background music. It offers manual fine-tuning control. User content will be automatically deleted from its servers within 24 hours. It operates on a pay-as-you-go model without subscriptions.

Read more about silenis.online

Users also considered
Aveni Assist logo

Automated meeting capture and compliance for advisers

learn more
Aveni Assist transcribes adviser–client meetings with high accuracy, applies speaker diarisation, and links content to CRM records. Compliance checks analyse transcripts for risks and maintain a searchable audit trail.

Read more about Aveni Assist

Users also considered
Listener logo

Speech to Text - fast and reliable

learn more
Listener is a product that transcribes speech to text in real-time. It supports multiple languages and domains and provides high accuracy, speech adaptation, timestamps, speaker diarization, and flexible model deployment.

Read more about Listener

Users also considered
Philips SpeechExec logo

Use the power of your voice with professional dictation

learn more
Philips SpeechExec Pro Dictation and Transcription Software is designed for authors to focus on recording with their preferred voice recorder, download dictations quickly, and automatically route to assistants or speech recognition to transcribe files.

Read more about Philips SpeechExec

Users also considered
AssemblyAI logo

Speech to text API with voice AI models

learn more
AssemblyAI provides speech-to-text transcription services through various API offerings, including pre-recorded and real-time transcription capabilities. The platform supports transcription in ninety-nine languages and includes features such as speaker identification, sentiment analysis, content moderation, and PII redaction. Additional functionality includes a Voice Agent API with turn detection and interruption handling, as well as an LLM Gateway for routing between different language models.

Read more about AssemblyAI

Users also considered
Voximal logo

A phone platform based on Asterisk propulsed by Voximal

learn more
Voximal is Asterisk's VoiceXML engine with state-of-the-art of latest text to speech and speech to text on-premise or online offer.

Read more about Voximal

Users also considered
Deepgram logo

Voice AI platform for speech recognition

learn more
Deepgram offers enterprise voice AI solutions via APIs for speech-to-text, text-to-speech, and voice agent capabilities. Its unified API includes language model orchestration, eliminating the need for separate tools. Features include real-time and batch processing, multilingual support in ten languages, speaker detection, sentiment analysis, intent detection, and topic identification for audio content.

Read more about Deepgram

Users also considered
GoVivace logo

Conversational AI and speech analytics solution

learn more
GoVivace is a conversational AI and speech analytics solution. It provides intelligent omnichannel chatbots and voice bots for businesses of all sizes.

Read more about GoVivace

Users also considered
ValueFlow logo

AI voice agents for conducting interviews

learn more
Interview agents that conduct voice-based interviews automatically. Simply share pre-configured interview links with your audience.

Read more about ValueFlow

Users also considered
Infercall logo

AI phone answering service for small businesses

learn more
Infercall is an AI-powered phone answering service that handles incoming calls for businesses around the clock. The platform trains on business information by extracting data from websites and uploaded documents, enabling it to answer questions about services, pricing, and availability. Infercall supports simultaneous call handling, appointment scheduling, lead qualification, call transfers, and provides full call transcripts and recordings with analytics.

Read more about Infercall

Users also considered
AICHE logo

AI-enabled software that transforms voice into text

learn more
AICHE transforms voice into polished text with one hotkey. Speak naturally - the AI delivers clean, structured output instantly copied to your clipboard. Available on Windows, Mac, Linux with privacy-first zero audio retention.

Read more about AICHE

Users also considered
AI-Powered Voice Assistants logo

Customer experience software for eCommerce businesses

learn more
AI-Powered Voice Assistants is a conversational marketing software that helps businesses recognize speech, interpret human language and optimize communications. Administrators can automate various repetitive tasks including insurance premium payment reminders and debt collection processes.

Read more about AI-Powered Voice Assistants

Users also considered
Gladia logo

Multilingual speech to text transcription API

learn more
Gladia provides an audio transcription API that converts speech to text through both asynchronous and real-time processing capabilities. The platform supports over one hundred languages and offers features including speaker diarization, sentiment analysis, named entity recognition, and word-level timestamps with sub-three-hundred-millisecond latency for real-time transcription.

Read more about Gladia

Users also considered
Amical logo

AI-based open-source speech-to-text application

learn more
Amical is an open-source speech-to-text application powered by generative AI technology that enables users to convert spoken words into text without using a keyboard. The application automatically understands context across different platforms, formatting dictation appropriately whether for professional emails or casual social media posts, while maintaining user privacy and delivering accurate transcriptions.

Read more about Amical

Users also considered