Mountain landscape

How to Transcribe Apple Voice Memos on a Mac

By Ryan Crabbe, developer of TotaLast verified against Tota on 27 July 2026

Voice Memos is where ideas go to be safe — and, too often, where they go to be forgotten. A walking brainstorm, an interview, a lecture, a client call recorded on your iPhone: getting it out of audio and into text is what makes it searchable, quotable, and usable. Here's how to do that on a Mac in about a minute, without uploading the recording anywhere.

TL;DR

Drag the recording out of the Voice Memos app onto your Desktop, drop the file into Tota's Audio Drop tab, and export the transcript as text, Markdown, SRT, JSON, or CSV. Everything runs on your Mac — no upload, no account, no minute limits.

Step 1: Get the Audio File Out of Voice Memos

Voice Memos stores recordings inside its own library, but exporting one is a single drag:

  1. Open Voice Memos on your Mac. If you record on your iPhone and use iCloud sync, your iPhone memos are already in the list.
  2. Click the recording you want, then drag it out of the window onto your Desktop (or into any Finder folder). macOS copies it out as an .m4a audio file.
  3. No iCloud sync? On your iPhone, open the memo, tap the share icon, and AirDrop it to your Mac instead.

Step 2: Drop It Into Tota

  1. Open Tota and go to the Audio Drop tab in the sidebar.
  2. Drag the .m4a file in (or click to pick it). MP3, WAV, AIFF, and other common formats work too.
  3. Watch the progress bar. Long recordings are processed in streaming chunks, so a two-hour lecture doesn't need a two-hour wait — and you can cancel at any point.

The transcription runs on the same Whisper-class models Tota uses for dictation, entirely on your Mac. You can pull the WiFi cable first if you want to check.

Step 3: Name the Speakers

If the memo has more than one voice — an interview, a meeting held over lunch — Tota detects the speakers automatically and gives each one a colour. Click a speaker label to rename "Speaker 1" to a real name, and the entire transcript plus every export updates. Timestamps are up to you: off, once per speaker turn (the default), or on every sentence.

Step 4: Export

Every transcript exports as plain text, Markdown, SRT subtitles, JSON, or CSV, each with copy and save options. Paste the minutes into a doc, drop the Markdown into Obsidian, or feed the JSON to a script — no converter needed.

What About Apple's Built-In Transcripts?

Honest note: on recent versions of macOS and iOS, the Voice Memos app can show a transcript of a recording itself, in supported languages. If all you need is to skim what you said and copy a paragraph, that built-in view may be enough — try it first.

Where a dedicated tool earns its keep:

  • Speaker labels. Apple's transcript is one undifferentiated block; Tota labels who said what and lets you rename speakers.
  • Export formats. Copying text is Apple's ceiling; Tota exports SRT for subtitling, CSV/JSON for pipelines, and Markdown for notes.
  • Any audio file. The same workflow handles MP3s from a recorder, WAV files from an audio interface, or a podcast episode — not just memos.

And if you're currently uploading recordings to a cloud service to get transcripts, the trade-offs are covered in Tota vs Otter.ai — the short version is that on-device transcription has no minute limits and leaves no copy of your recording on a server.

The Bigger Picture

Voice memos are just one input. The same Audio Drop workflow covers meeting recordings, research interviews, and subtitle generation — the full picture, including how on-device transcription compares to Otter.ai and MacWhisper, is on our transcribe audio files on Mac page, and the feature reference lives in the Audio Drop documentation.

Frequently asked questions

How do I transcribe a voice memo recorded on my iPhone?

If Voice Memos syncs via iCloud, the recording already appears in the Voice Memos app on your Mac — drag it out to the Desktop and drop it into Tota's Audio Drop. If you don't use iCloud sync, AirDrop the memo from your iPhone to your Mac instead; it arrives as an M4A file that Audio Drop accepts directly.

Is there a length limit on the recording?

No. Tota processes long files in streaming chunks, so multi-hour recordings work, and there are no per-minute quotas or monthly caps — unlike cloud services such as Otter.ai, which meters free accounts at 300 minutes per month.

Does the voice memo get uploaded anywhere?

No. Tota transcribes the file entirely on your Mac — speech recognition and speaker identification both run locally, and the transcription works with WiFi turned off. Nothing is uploaded, so there is no copy of the recording on anyone's server.

Can it tell different speakers apart in a voice memo?

Yes. Tota detects speakers automatically and labels each turn. Click a label to rename "Speaker 1" to a real name and the whole transcript and every export update. If the voices in a recording genuinely can't be separated, you get a clean single-speaker transcript rather than a garbled guess.

What file formats does Audio Drop accept besides voice memos?

Voice memos are M4A files, and Audio Drop also accepts MP3, WAV, AIFF, and other common audio formats — so the same workflow covers meeting recordings, interviews, and podcast audio.