Fully local processing
Transcribes audio locally so files stay on your computer, which the README presents as the product’s main privacy and security benefit.
EchoTranscribe is an open-source desktop app for local audio transcription using Whisper models. It keeps files on-device, supports batch transcription, and exports transcripts as TXT, SRT, or JSON.
EchoTranscribe is an open-source desktop app for audio transcription built on local AI. It uses Whisper models to convert speech to text while keeping audio files on the user’s machine, which positions it as a privacy-focused option for offline transcription workflows.
The app is packaged as a cross-platform Tauri desktop application with a Python backend and a React/TypeScript frontend. The README presents it as suitable for single-file transcription, batch processing, transcript review, and export to common formats such as TXT, SRT, and JSON.
Transcribes audio locally so files stay on your computer, which the README presents as the product’s main privacy and security benefit.
Uses Whisper-based models for transcription, with model-size choices such as Tiny, Base, Small, and Medium described in the README.
Handles multiple audio files in one session and supports batch transcription with per-file progress tracking.
Supports MP3, WAV, FLAC, M4A, OGG, and WebM input files, with the README listing a 500 MB max size for each.
Shows word-level timestamps to help users navigate transcripts and review specific parts of the audio.
Exports results to TXT, SRT, or JSON and allows both individual and batch export flows.
Use EchoTranscribe to turn a single recording into text while keeping the audio file local on your machine.
Process several recordings in one run and review progress file by file, which fits recurring transcription workloads.
Review transcripts with word-level timestamps when you need to jump back to a specific section of the audio.
Export the finished transcript in TXT, SRT, or JSON depending on whether you need plain text, subtitle-style output, or structured data.
Run the app on Windows, macOS, or Linux for local transcription without relying on a hosted service.
Yes. The README says you can run the app on Windows, macOS, and Linux.
The README shows TXT, SRT, and JSON as export options, both for individual files and batch exports.
The quick start uses Node.js, Python, and Rust for development, and the backend runs locally with the Tauri desktop app.
The product is aimed at local transcription work. The source emphasizes privacy, secure local processing, and support for batch transcription and word-level timestamps.
No integrations are documented in the provided sources. The available documentation focuses on local file transcription and export rather than third-party connectors.
QuickQuill is a macOS dictation and transcription app that runs locally on the device. It helps users record meetings, transcribe audio, generate summaries, and export notes without using a cloud service.
Speech to Text Converter is a browser-based transcription tool for live dictation and uploaded audio or video files. It offers a free tier for short tasks and a Pro plan for unlimited transcription, AI summaries, translation, speaker identification, and advanced exports.
Pewbeam is a church presentation app that listens to sermons, detects Bible verse references in real time, and displays the matching passage on screen. It is built for pastors, projection teams, and church media volunteers who want to reduce manual slide control during live services.
Dictato is a Mac dictation app that transcribes speech into text in any app using an on-device, offline workflow. It supports multiple transcription engines, optional cleanup and translation, and a one-time purchase license.
Sanota is an app that turns spoken memories, reflections, and interviews into clear written stories. It supports personal storytelling, family history, and shared memories, with guided prompts and subscription pricing.
Carbon Voice is an asynchronous voice messaging app for teams and individuals, with transcripts, AI catch-up, and cross-device access. It helps people and agents communicate without needing a live call.