A practical guide · ~6 min read
Turn Voice Memos Into AI Memory
The best thinking most people do is in the car, on a walk, or 30 seconds after waking up. Almost none of it gets written down — you know that when you finally sit at a keyboard the sentence is gone. Voice memos catch it. Which is why the folder on your phone quietly holds better material than your journal. Here's how to get it out.
Why voice is the highest-signal capture format
Typed thoughts are edited before they leave your fingers. Spoken thoughts aren't. You hesitate, backtrack, land on a phrasing you'd never have written but that turns out to be the actual truth. If you've ever transcribed one of your own voice memos and thought "I said that?" — you've noticed the effect.
For AI memory purposes, that's gold. A persona built from your typed writing sounds like your polished self. A persona built partly from voice sounds like the version of you that's still figuring it out in real time — which is almost always the more interesting version.
Getting memos off an iPhone
Apple deliberately makes this slightly awkward. The Voice Memos app won't hand a file directly to Safari, so you have to save through Files first.
- Open Voice Memos.
- Tap the memo → the three-dot menu → Share.
- Choose Save to Files → pick a folder in On My iPhone or iCloud Drive.
- Optional but faster: instead of Save to Files, AirDrop to your Mac. The .m4a lands in Downloads.
To batch multiple memos, tap Edit in the memo list, select several, then Share. You can grab a month's memos in one action.
Transcription: what to expect
Modern transcription is 95%+ accurate on clear speech. The failure modes are predictable: proper nouns (names of people, companies, places) come through phonetically, and if you're recording in a moving car with wind noise the model sometimes hears words that were never there. Skim the transcript once before you trust it as memory.
Konshus uses ElevenLabs Scribe for transcription — audio is processed and deleted, not stored, not used for training. Any decent tool will document this on a privacy page. If a transcription service can't tell you what happens to the raw audio, don't use it for voice memos.
Then: memory, not transcripts
A folder of transcripts is only marginally better than a folder of audio — you still can't skim it. What makes voice memos useful long-term is a tool that clusters them into themes: "you keep circling the same ambition question," "your Sunday morning walk memos are always about your parents." That's memory. A transcript is just text with worse formatting.
Konshus imports voice memos and does that clustering automatically — each memo becomes a voice_memo artifact and the recurring ideas surface as atoms your Konshus can quote back to you.