VoiceMemo records where work actually happens — meetings, classes, calls, passing thoughts — and turns every conversation into a clean record: who said what, what was decided, what happens next. Speaker-separated, searchable, and private by default.
Multi-person recordings come back split by voice — every paragraph labeled and timestamped, up to six speakers. Tap a label to give it a real name; the AI reads self-introductions and how people address each other, then suggests who's who. You confirm, and names carry into summaries and PDF exports.
Suggestions come only from the conversation itself. Nothing is applied without you.
After each recording, VoiceMemo spots the names and terms generic engines get wrong — your colleagues, your jargon, your product names. Confirm a correction once, and every future transcript writes it properly. Accuracy that compounds, in a vocabulary that's yours alone.
People and terms live in one directory — it also powers speaker-name suggestions and grounds your summaries.
Ask in plain language. VoiceMemo reads across every recording you've made and answers with the exact spot it found it. Your conversations stop being files — they become a memory that talks back.
Bind the Action Button, Lock Screen, or Siri. One press and you're recording — pocket, screen off, mid-walk.
Overview, key decisions, action items — pulled out of your ramble in seconds, listed separately.
Select a passage, note what to change, keep adding notes — then apply them all in one pass.
Lock the screen, switch apps, lose signal — transcription carries on in the background and retries on its own.
PDF for the record, an image card for the group chat, plain text, or the original audio. Format first, then scope.
Transcribed on your device by default. Audio lives on your phone and your own iCloud — never on our servers.





The capabilities that sell those devices — speaker separation, accurate transcription, structured summaries — are software. VoiceMemo puts them in the phone you already carry. The hardware only earns its price for phone-call recording and wearable capture.
Read the honest comparison →Processed entirely on your iPhone. Neither audio nor text is uploaded — it works even in airplane mode.
Higher accuracy on long recordings, with speaker separation for up to six voices. Used only when you choose it, over an encrypted connection. Runs on your monthly AI minutes — 60 free, 1,200 with Pro.
An AI minute = one minute of your recording being processed by cloud AI — charged once per recording (cloud transcription or its first summary); re-generating and revising after that are free.
Prices in USD, billed through the App Store · manage or cancel anytime in Settings
Free to start, no account. The transcript, the speakers, the decisions, the to-dos — all ready before you're back at your desk.