Skip to content
Trustample

Convert M4A to text

Upload the .m4a straight from your phone or Mac, with no converting to MP3 first, and get clean, editable text in seconds. Free to start.

Free: 60 min per 30 days, files up to 30 min / 50 MB · Paid plans: 10 to 100 hrs, files up to 2 GB. Big video? Upload just its audio track: same transcript, much faster upload.

Encrypted in transit · audio & video auto-delete in 30 days · never used to train AI

Optional settings

Naming the language beats leaving this on auto. If the recording has more than one, pick the one spoken most.

Get a second copy translated into another language (Pro).

Uncommon words the AI might misspell. List them so they come out right.

Last updated 6 August 2026

Skip the step most people waste time on: you do not need to convert an M4A into MP3 or WAV before transcribing it. M4A is just AAC audio in an MP4 container, and speech recognition reads it directly. Any guide telling you to convert the file first is adding a lossy, pointless detour. Drop the .m4a in as it is.

M4A is everywhere because Apple made it the default: every iPhone and iPad Voice Memo is an M4A, QuickTime audio recordings are M4A, Zoom saves its local audio-only track as M4A, and plenty of Android recorders produce them too. If you recorded a meeting, lecture, interview or a note to yourself in the last few years, an .m4a is probably what you're holding.

We converted an M4A to MP3 to see what it costs you

The advice to convert first is everywhere, so in August 2026 we tested it against our own live pipeline. We recorded a 27 second spoken memo (names, dates, a couple of money figures and some office jargon), encoded it as AAC in an M4A the way a phone does, then re-encoded that same M4A to MP3 at the same bitrate, which is exactly what conversion guides tell you to do. Both files went through the identical transcription request.

The M4A returned the surname correctly: Priya Raghunathan. The MP3 made from it returned Priya Ranganathan. Everything else matched across both runs, including March 14, April 2, 42,000, $65,000, SOC 2 and a second surname, Delacroix. The converted file also broke one sentence in a different place. It was not even smaller: 432 KB against the original's 449 KB, so the detour bought no disk space either.

Two things are worth taking from that. The first is that conversion cannot add detail that was never captured, it can only discard some, so the original file is always the safer upload. The second is more useful: both runs came back with 99.9 percent confidence, including the one that got the name wrong. A confidence score will not warn you about this, which is why the two-minute check below matters more than any number the engine reports.

Being straight about the limits of that test: it was one sample and one word of difference, not a benchmark. It points in the direction the theory already predicted rather than proving a rule.

Getting the .m4a off your phone or Mac

On iPhone or iPad, open Voice Memos, tap the recording, tap the share icon, and either save it to Files, AirDrop it to a Mac, or email it to yourself. You can also skip the transfer completely: open this page in Safari on the phone and upload straight from Files, which turns your phone into the whole workflow rather than the first step of one.

On a Mac, QuickTime audio recordings and Voice Memos both land as .m4a and drag straight from Finder into the upload box. Worth knowing: macOS dictation cannot help here, because it types what it hears through the microphone live and has no way to open a file you already recorded. The same is true of Windows voice typing. Uploading the file is not the workaround, it is the actual answer.

Zoom is the other common source. Its local recordings produce an audio-only track saved as M4A alongside the video, and that audio file is the better upload: same speech, much smaller file, much faster transfer.

The two-minute check that catches most errors

Speech recognition is strongest on ordinary conversational words and weakest on proper nouns, which is precisely what our test showed. Names are where corrections cluster, so read those first. If a name matters and appears throughout the recording, fixing it once in the synced editor and then searching the transcript for the other spellings takes less time than reading the whole thing twice.

Better still, prevent it. Type recurring names, company names and specialist terms into the vocabulary box before uploading, a field available on every plan including free, so the engine expects those words instead of guessing at them.

Numbers deserve the second glance. Spoken figures are converted to digits automatically, which is what you want in a memo you will act on: our test turned March fourteenth into March 14 and sixty five thousand dollars into $65,000 without being asked. That is helpful right up until a currency amount or a percentage is the one thing that has to be exact, so check those against the audio before the transcript becomes a decision.

Example transcript

A sample of the output. Every line carries a timestamp and a speaker label, and in the editor each word links back to that exact moment in the audio, so checking a quote takes seconds.

Example output
00:00Speaker 1Thanks for joining. Let's start with the quarterly numbers before we get to the roadmap.
00:07Speaker 2Sure. Revenue was up eleven percent, mostly from the new self-serve plan.
00:14Speaker 1And churn?
00:16Speaker 2Down about half a point. The onboarding changes seem to be helping.

How it works

  1. 1Drop the .m4a straight in from Files, Finder, or AirDrop. No conversion needed.
  2. 2The AI transcribes it with punctuation, usually in seconds for a short memo.
  3. 3Click any word to fix it, then copy the text or export a document.

Frequently Asked Questions (FAQs)

Do I need to convert M4A to MP3 first?

No, and our own test suggests converting can cost you. M4A is a container holding AAC audio, which the engine reads natively. When we re-encoded one M4A to MP3 and ran both, the original got a surname right that the converted copy misheard, and the MP3 was not meaningfully smaller. Re-encoding is lossy by definition: it cannot add anything back. Upload the original.

Is transcribing an M4A less accurate than an MP3 or WAV?

No. Accuracy is set by the recording, not the container: microphone distance, background noise, crosstalk and how clearly people speak. A clean M4A beats a noisy WAV every time. The one thing that does move the needle is giving the engine your vocabulary (names, jargon) before it starts.

Can I transcribe M4A to text for free?

Yes. The free tier is a monthly hour of memos in files up to 30 minutes, with plain-text export, no card, and the same engine and synced editor paid users get. It renews every 30 days. Files over 30 minutes, or exports beyond plain text, need Basic ($12/mo) or Pro ($19/mo).

How do I get a Voice Memo off my iPhone?

In the Voice Memos app, tap the memo → share icon → save to Files, AirDrop it to your Mac, or email it to yourself. Then upload the .m4a. You can also skip the transfer entirely: open Trustample in Safari on the phone and upload straight from Files. Safari handles the whole flow, so your phone gains a transcript, not another app.

How do I transcribe an M4A on a Mac?

Upload it in any browser; macOS dictation types what you speak into the microphone, so it can't take an existing file. Drag the .m4a from Finder into the upload box and the transcript comes back in the same tab. Same for Windows.

My memo is a conversation. Can it label who's speaking?

Yes, on every paid plan: when the recording has more than one voice, segments arrive labeled Speaker 1, Speaker 2 and so on, and you can rename them; the label updates through the whole transcript. Talk-time stats come with the labels; the AI-written speaker insights summary is the Pro extra.

What if my M4A is a long meeting or lecture?

Free covers files up to 30 minutes. Basic handles up to 3 hours, Pro up to 6 hr 40 min or 2 GB. If you're on free with a 60-minute recording, splitting the file in two with any free audio editor works, or upgrade, which is what the plans are for.

Are voice memos private after upload?

Yes, and memos deserve it: they are often half-formed thoughts you would never say publicly. Storage is encrypted, recordings are never used to train AI models, the audio auto-deletes after 30 days (or instantly if you delete it yourself), and paranoid mode erases the memo the moment its text exists.

Worth knowing: Our conversion test was a single 27 second sample with one word of difference between the two files, so treat it as a spot check rather than a benchmark. It also used clean, studio-quality speech: a memo recorded across a noisy room or with a phone in your pocket will lose accuracy for reasons that have nothing to do with the file format. Test your own recordings on the free plan before committing a long one.

Related tools

See all converters on the transcription tools page.

Browse all transcription tools.

Chat with us