Drag in audio or video and get an accurate, speaker-labelled transcript — running OpenAI's Whisper models entirely on your Mac.

MacWhisper wraps OpenAI's Whisper speech-recognition models in a clean native app. Drop in a file — or record a call, a meeting, or system audio — and get a transcript you can search, edit, and export, without anything leaving your machine.
Local transcription used to mean the command line. MacWhisper makes it a two-second drag-and-drop, with a side-by-side editor, speaker detection, and translation built in. It is fast on Apple silicon, it handles dozens of languages, and the privacy story is simple: the audio never touches a server.
On-device Whisper models, from tiny to large-v3, with Apple silicon acceleration
System-audio and microphone recording, including automatic meeting capture
Speaker recognition that labels who said what
Built-in editor with search-and-replace and word-level timestamps
Export to SRT, VTT, CSV, plain text, and Markdown
AI prompts for summaries and translations of any transcript
The free tier covers smaller models and short files. A one-time Pro license (around $59) unlocks the largest models, batch processing, recording, and the AI features.