The workflow is record, combine, then write: record chapters or scenes out loud as separate Voice Memos, get them onto your Mac, transcribe and combine them into one ordered Master Transcript, then bring that transcript into a writing app or an AI assistant as raw source material. The transcript is not a finished draft. It is the material you shape into one, the same way a rough outline or a stack of notes would be.
Voice Memo Exporter does this in one pass: every memo into one searchable Master Transcript. $49 early access, runs privately on your Mac.
Books written from voice tend to work best when each recording maps to something small: one scene, one chapter beat, one argument, not an entire book in a single hour-long ramble.
1
Record one idea, scene, or chapter section per memo
Shorter recordings are easier to transcribe accurately and easier to find later. A 10 to 20 minute memo per chapter section is a reasonable target.
2
Say what you're recording at the start
A spoken line like 'Chapter 4, the part about the move to Osaka' at the start of a memo becomes a natural heading once transcribed, and it helps you find the right recording months later.
3
Rename memos as you go if you can
Apple's default titles are unhelpful (New Recording 213). A few seconds renaming each one in the Voice Memos app pays off later when you are trying to find a specific chapter.
Get everything onto your Mac and combined
A book is not one recording, it's dozens or hundreds. The value comes from having them all in one ordered document, not scattered across an app's scrolling list.
1
Sync or export recordings to your Mac
Turn on Voice Memos in iCloud settings on the iPhone and sign into the same Apple Account on the Mac, or export recordings manually as .m4a files.
2
Transcribe each recording
Apple's built-in transcripts (macOS Sequoia and later) work per recording. For a whole manuscript's worth of memos, a bulk local transcription pass saves the copy-paste ritual.
3
Combine transcripts into one ordered document
Order matters for a book. A single Master Transcript with recordings in date or chapter order gives you one file to scroll, search, and hand off, instead of a folder of loose text files.
The honest cost
What manual transcription and stitching actually costs
Apple's per-recording transcripts require selecting, copying, and pasting each one into a document by hand, then manually ordering them into chapters. For a 200-recording manuscript, that is realistically several hours of copy-paste before you have written a single sentence of actual prose.
No batch transcription or export in Apple's Voice Memos app.
Ordering and combining transcripts into chapters is manual work.
Hours spent assembling material is time not spent writing.
The transcript is raw material, not a draft
This is worth saying plainly: a transcript of you talking is not a finished chapter. Spoken language rambles, repeats itself, and trails off in ways prose does not. Local transcription is also best-effort, so expect rough patches, especially with background noise or fast speech. Treat the Master Transcript as ore you refine, not ore you publish. Read it, pull out the good lines and the real structure, and rewrite from there, whether by hand or with an AI assistant helping you find the shape inside a hundred pages of your own talking.
There's a faster way than doing this by hand.
Voice Memo Exporter does the steps above in one pass: every Apple Voice Memo on your Mac exported, transcribed locally, and combined into one searchable Master Transcript. $49 early access, private by design.
Once you have a combined Master Transcript, a language model can help you find the book inside it: grouping recordings by theme, flagging places where you repeat the same idea across different sessions, or drafting a chapter outline from the raw material. The output is only as good as the input, so a well-ordered, complete transcript matters more than a clever prompt. Paste the Markdown version into Claude or ChatGPT and start with something direct, like asking it to summarize the recurring themes across the whole file before you ask it to draft anything.
A note from the founder
I'm writing my own first book this way: talking out chapters into Voice Memos, combining them into a Master Transcript, then rewriting from that material by hand and with AI. It works, but it is genuinely a rewrite, not a paste job. I'll be sharing more about that specific workflow separately as the book gets closer to done.
Related questions
Questions about turn voice memos into a book.
Can I write a whole book just from voice memos?
You can generate the raw material that way, but expect to rewrite. Spoken language needs editing to read as prose. The transcript gives you honest first material, themes, and structure, not a publishable draft.
How accurate does the transcription need to be for a book draft?
Accurate enough to recognize what you meant when you reread it. Local transcription is best-effort, so plan to review the transcript against the audio for any passage you intend to use closely, rather than trusting it word for word.
Should I record one long memo per writing session or many short ones?
Many short ones tied to a specific scene or idea are easier to transcribe accurately, easier to find later, and easier to slot into an outline than a single sprawling recording.
Can I use ChatGPT or Claude to help structure the book?
Yes. Once your recordings are combined into one Master Transcript, pasting the Markdown version into an AI assistant lets it work from your actual words and recurring themes, which produces more recognizable output than asking it to invent content from a prompt.
Does Voice Memo Exporter write the book for me?
No. It builds the archive and the combined Master Transcript from your recordings. Writing, rewriting, and structuring the book is still your work, with or without AI assistance.
Or skip the manual version entirely.
Voice Memo Exporter runs this whole workflow in one pass on your Mac: every memo exported, transcribed locally, and combined into one searchable Master Transcript you own.