If you've ever opened the Voice Memos app on an iPhone and recorded something, you have an M4A file. It's the default Apple format — AAC audio in an MP4 container, optimised for the iPhone's mic — and most online transcribers force you to convert it to MP3 first. MDisBetter accepts M4A directly: drop the file from your iCloud or AirDrop, get Markdown back.
Why M4A matters for mobile-first workflows
The vast majority of voice content in 2026 is captured on phones, and the vast majority of phones default to M4A or its near-cousin AAC. iPhone Voice Memos = M4A. Most podcast apps that record interviews on iPhone = M4A. iOS Music app rips = M4A. Demanding "convert to MP3 first" adds a step that loses quality (re-encoding lossy audio is always slightly worse than the original) and breaks momentum. Direct M4A support means the file goes from your iPhone to Markdown without a round-trip through Audacity or an online converter.
Common iPhone-recording use cases
Field interviews: journalist hits Record on the iPhone, conducts a 45-minute interview, AirDrops the M4A to laptop, drops it into MDisBetter, has a structured transcript with headings and timestamps in 2 minutes. Meeting capture without a bot: place the iPhone on the conference table, hit Record, end the meeting, transcribe the M4A — no calendar invite to a meeting bot, no recording disclosure beyond what your jurisdiction requires you to handle yourself. Voice notes for self: rambling 5-minute idea dumps while walking, transcribed into searchable Markdown notes for later review.
Structure of the output
iPhone Voice Memos are typically captured in noisy environments with one or two voices in close proximity. The transcript comes back as continuous punctuated text under a top heading, with ## H2 sections where the topic shifts. We do not identify or label speakers, so a two-person conversation reads as one flowing transcript; for very noisy environments or 3+ overlapping voices, word accuracy degrades and you may need to spot-check passages. Inline timestamps mark section boundaries so you can jump back to the audio if something needs review.
Before / After
Before (PDF):
[M4A audio file]
AAC audio in MP4 container, 4.2 MB, 18 minutes mono
Metadata: iPhone Voice Memo, recorded 2026-05-09, location tagged
(no extractable text — pure audio waveform)
After (Markdown):
# Voice Memo — May 9, 2026
## Opening Notes
[00:00:01] Recording the meeting with the contractor. Let me confirm, you can hear me OK on this end? Loud and clear, yeah.
## Scope Discussion
[00:00:14] OK so the scope as I understand it is the kitchen rebuild plus the bathroom. Is that right? Right, kitchen and the master bath. Not the second bath…
Frequently asked questions
How do I get the M4A file off my iPhone to upload?
Three options. AirDrop from Voice Memos to your Mac (one tap, fastest). iCloud Drive: in Voice Memos, share to Files, save to iCloud, access from any device. Email to yourself: works for files under ~25 MB; longer recordings may exceed mail attachment limits. Once the M4A is on a laptop, drop it into the MDisBetter web tool.
Are the two people in an iPhone interview labelled separately?
No. We do not identify speakers, so a two-person interview comes back as one continuous transcript rather than alternating labelled turns. What you do get is accurate punctuated text, ## H2 sections at topic shifts and inline timestamps for jumping back to the audio. If you need who-said-what attribution, run the file through a dedicated diarisation tool such as WhisperX or pyannote alongside our Markdown.
What about iPhone-recorded calls or FaceTime recordings?
Recorded phone calls saved as M4A (via screen recording or third-party call-recording apps) work the same as any M4A: you get the spoken content as punctuated Markdown, without speaker attribution. Note that recording calls without consent is illegal in many jurisdictions — that's on you, not the converter.
Are M4A files from Voice Memos lossless?
No — M4A from iPhone Voice Memos uses lossy AAC encoding (similar quality tier to MP3 at 128 kbps but more efficient). For most spoken-word use cases the loss is inaudible. If you need lossless capture, switch the iPhone Voice Memos setting to "Lossless" (introduced in iOS 15) which captures uncompressed PCM — file sizes go up roughly 10×.
Can I batch-convert multiple M4A files from a recording session?
The MDisBetter web tool processes one file at a time. For a multi-file recording session (e.g. a day of field interviews captured as separate M4As), upload each one in turn — typical processing time is a couple of minutes per 30-minute file. For automated batch pipelines, faster-whisper running locally can process a folder of M4As in a script loop without per-file interaction.