Guide
Transcribe and Summarize Recordings on iPhone and Mac — Without the Cloud
You have a recording — a meeting, a lecture, an interview, a voice memo, maybe a video — and you want it as text you can search, quote, and ask questions about. Most tools that offer this route your audio through someone else's servers. This guide covers the alternative: transcription that runs entirely on your iPhone, iPad, or Mac, what Apple already gives you for free, and what to do once the transcript exists.
Why recordings deserve the off-cloud default
A recording is different from a document you wrote. It contains other people's voices — things they said in a specific room, to a specific audience, with specific expectations. A weekly standup covered by an NDA. A user-research interview conducted under a consent form that named who would hear it. A client call. A lecture where classmates asked questions they did not expect to be uploaded anywhere.
When you send that file to a cloud transcription service, you are making a data-handling decision on behalf of everyone in the recording. Sometimes that is fine — a public podcast has no confidentiality to protect. But for meetings, interviews, and anything work- or study-related, the safe default is the opposite: the audio stays on the device that holds it, and the transcription comes to the audio rather than the audio going to the transcription.
There is also a mundane practical upside: no upload step means large files (a two-hour lecture video, say) transcribe without a network connection at all — on a plane, in a basement lecture hall, on hospital Wi-Fi you do not trust.
What your iPhone and Mac already do — use it if it is enough
Before installing anything, know what Apple ships natively, because for many people it fully answers the question:
- Voice Memos can transcribe recordings on supported devices. Open a recording, view the transcript, copy it out. The transcription is generated on the device.
- Notes can record audio inside a note and attach a transcript, and phone-call recordings land in Notes with a transcript. On Apple Intelligence devices you can also get a summary of the transcript.
If your need is "I recorded a memo and want its text," stop here — Voice Memos is the answer, it is free, and it is already on your phone. This page will not pretend otherwise.
The native tools stop at three specific walls:
- They transcribe what they record. A file that already exists — an exported Zoom recording, a lecture video, an MP3 a colleague sent, years of old interviews — has no clean path into Voice Memos or Notes for transcription.
- Each transcript lives inside its one recording. You can find a phrase, but you cannot ask a question — there is no way to ask something across twenty recordings and get an answer that tells you which recording, and where, it came from.
- No citations, no restraint. A native summary gives you a paragraph; it does not point at the exact passage it came from, and it will not tell you when the recording simply does not contain what you are looking for.
Can ChatGPT summarize an audio recording? The honest answer
Yes. Cloud AI assistants can transcribe and summarize audio, and generally do it well. If capability were the only question, this page would be shorter.
The real question is routing. To summarize your recording, a cloud service has to receive it: the file is uploaded and processed on the provider's servers, under whatever retention, logging, and training policies apply at that moment — policies you would need to read, and re-check, because they change. None of that requires assuming any provider is careless. The point is simpler: the confidentiality decision is made the instant the file leaves your device, before any policy on the other end matters. For an NDA-covered meeting or a recorded patient conversation, the upload itself can be the violation, regardless of what happens to the file afterward.
So the honest framing is: use cloud AI for audio that has no confidentiality to protect, and use on-device transcription for audio that does. (If you want to know how to check what any AI app actually collects — rather than taking its marketing's word for it — see what "Data Not Collected" actually means.)
Transcribing an existing audio or video file on-device, step by step
The gap the native tools leave — import a file you already have, transcribe it locally, then actually work with the text — is what OpenIntelligence is built for. It is free to download on the App Store, with one requirement stated plainly up front: it needs iOS, iPadOS, or macOS 26 or later on Apple Intelligence-capable hardware — an iPhone 15 Pro or newer, or an iPad or Mac with an M1 chip or newer. On older devices it will not run, and the native options above are the honest recommendation instead.
- Install the app and create a collection. A collection is just a bucket — "CS 301 Fall," "Acme project," "Interviews."
- Add the recording. From the Files app or the share sheet on iPhone and iPad, or by dragging files in on the Mac. Audio and video both work — for a video, the soundtrack is what gets transcribed.
- Let it transcribe locally. Speech recognition runs on the device using Apple's on-device transcription. Nothing is uploaded — there is no account and no API key to even send anything to. If you want proof rather than a promise, turn on airplane mode first; the transcription proceeds anyway. How long it takes depends on the recording's length and your hardware.
- The transcript is indexed alongside your documents. The text goes into the same local index as any PDFs, slides, or notes in the collection — full-text search via SQLite FTS5 plus semantic retrieval, accelerated across the CPU, GPU, and Neural Engine.
- Ask questions. Answers come with tappable citations pointing at the exact transcript passage they came from. And when your recordings do not contain the answer, the app abstains and says so — it does not fill the gap with a guess. If you just want the raw transcript, that is there too.
After the transcript: search a semester, not a file
A transcript on its own is a wall of text. The useful version of this workflow starts when recordings and documents share one index.
The clearest case is a student's: put twelve weeks of lecture recordings and the professor's slide decks into one collection. Now "when did we cover eigenvalues, and what did the slides add that the lecture skipped?" is a question you can actually ask — and the answer cites the specific lecture's transcript segment and the specific slide page, so you can jump to either and verify. Scanned handouts work too; printed pages are read with on-device OCR before indexing.
The same shape works for meetings: the recording of the negotiation next to the contract being negotiated, or a quarter of project calls next to the spec they kept referencing. The mechanics of asking questions across many documents at once — and how the citations and abstention behave — are covered in searching inside multiple PDFs at once, and the broader picture of what local models can and cannot do is in what on-device AI can actually do.
What this is not
To save you a download if you are looking for something else:
- It is not a live meeting bot. It does not join Zoom, Meet, or Teams calls, does not record anything itself, and does not caption a meeting in real time. You record with whatever you already use — Voice Memos, your meeting tool's own recorder — and bring the file in afterward.
- It is not a study-tools suite. No flashcards, no quiz generation, no spaced repetition. It answers questions about your material with citations; turning that into study artifacts is up to you.
- It has no calendar integration. It will not attach a transcript to the meeting it came from or know your schedule. What it knows about a recording is its name and its contents.
One more honest limit that applies to every transcriber, local or cloud: quality tracks the audio. A phone in the middle of a conference table, heavy crosstalk, or a bad microphone will degrade any transcript, and Apple's on-device models are no exception. Skim the transcript before you quote someone from it.
Requirements and limits
- System: iOS 26, iPadOS 26, or macOS 26 or later.
- Hardware: Apple Intelligence-capable devices only — iPhone 15 Pro or later, or an iPad or Mac with M1 or newer. This is a hard floor, not a soft recommendation; the on-device models require it.
- Price: free to download. A one-time Lifetime unlock and Pro monthly/annual subscriptions exist; prices vary by territory, so check the App Store listing for yours.
- Network: none required. Transcription, indexing, retrieval, and answering all run on the device and work in airplane mode. No account, no API keys.
- Privacy: the App Store privacy label is "Data Not Collected." Here is how to verify that yourself — for this app or any other.
If the native transcript in Voice Memos covers your need, use it. If your recordings are files you already have, or the transcript is the beginning of the work rather than the end of it, that is the case this workflow — and this app — exists for.