Voice Studio

How to record a podcast episode on iPhone

A podcast episode and a two-minute voice memo are technically the same action — press record, talk, press stop — but almost everything around that action is different once the recording is meant to be edited and published rather than played back once by the person who made it.

The audio itself does not need anything special. A phone microphone is entirely capable of capturing a clean solo voice for forty minutes the same as it is for four. What changes with a podcast is what happens on either side of the recording: you generally want a setting that will not need re-deciding halfway through a long take, a file that a separate editing tool can actually open, and a session that survives an hour uninterrupted rather than needing to survive thirty seconds.

The quality setting: high is the sensible default, not ultra

It is tempting to reach for the highest number on the assumption that a podcast, being a thing other people will listen to, deserves the best setting available. For speech, that assumption costs you file size and buys you nothing you can hear. A microphone-close solo voice is well within what a mid-tier sample rate captures cleanly — the ceiling on speech clarity is set by the room and the microphone distance, not by which of the available quality tiers was selected. Ultra mainly matters if you have a separate reason to want the highest-fidelity file possible; it is not a "sounds better on the podcast" setting.

The practical reason to care either way is length. An hour of solo talking at a higher sample rate is a meaningfully bigger file than the same hour at a lower one, and that difference compounds across a season of episodes sitting on the phone waiting to be edited. Picking a mid-tier setting once, rather than defaulting to the top one out of caution, is worth doing before you record the first episode rather than after the twentieth.

Where the phone sits still matters more than any setting

A quality setting changes how finely the microphone signal is sampled, not what the microphone actually picked up. Distance from your mouth, a room with hard surfaces, a fan or a laptop nearby — all of that reaches the recording untouched by which tier you chose, and none of it is fixable afterwards the way a quiet room and a close phone prevent it in the first place. If background noise is a recurring problem in the space you record in, it is worth reading through separately, because the fix lives in placement and room choice, not in the recording settings.

Surviving a solo session that runs long

A podcast episode is exactly the kind of recording where the ordinary threats to a long session actually show up, because a voice memo rarely runs an hour and an episode often does.

None of this is specific to podcasting — it is the same handful of things that threaten any long recording — but a podcast is where they are most likely to actually cost you something, because re-recording an hour of solo talking to match the energy of the first take is a genuinely different problem from re-recording a two-minute note.

Getting the file into an editor afterward

An iPhone voice recorder set to a sensible quality level hands you a compressed .m4a file — an AAC recording, not a raw or uncompressed one. That is not a limitation you need to work around: the file opens in any audio editor, drops onto any timeline, and carries a solo voice recording with essentially nothing audibly lost to compression. If a particular editing pipeline insists on an uncompressed format for its processing chain, converting an .m4a into that format is a one-step job in the editor itself — the direction that loses nothing is going from compressed to uncompressed, not the other way around.

What matters more than the format, in practice, is getting the file off the phone before you need it. Sharing it straight out through the share sheet, or pulling it in bulk through the Files app, both work; a backup export is a separate thing meant for restoring everything at once, not for handing one episode to an editor.

A transcript is worth having, even for one voice

A transcript of a solo episode is a fast way to write show notes without re-listening to the whole thing, and it is searchable in a way the audio is not — useful the moment you cannot remember which episode you told a particular story in. It is worth knowing one limit ahead of time: speech recognition on iPhone does not label speakers, so if an episode ever has a guest, the transcript comes back as one continuous block of text with no indication of who said which line. For a solo show that limit never comes up; for an interview-format one, it is worth reading about separately before you plan around it.

Export is TXT or JSON — plain text for pasting into show notes, structured JSON if you are feeding the transcript into something else. Neither is a subtitle file with timing built in, so if an episode needs captions for video, the transcript is a starting point for that work rather than a finished file.

How Voice Studio fits into this

Voice Studio records mono AAC in .m4a at your choice of three sample rates, with no limit on how long a single recording or how many recordings you keep — an hour-long episode is not a special case, and neither is a full season sitting on the phone. Transcription attempts the on-device path first and only retries against Apple’s server if that local pass comes back empty or errors; the transcript is editable afterward, which matters for the names and terms a recognizer reliably gets wrong. A call interruption pauses the recording and resumes into the same file rather than ending it, and background audio means the screen locking or the app losing focus does not stop a take either.

Two things worth knowing sit behind Pro rather than the free app: noise reduction paired with live monitoring, for a room you cannot fully control and want to hear as you record rather than discover afterward, and a fixed recording length, for anyone who records to a set time slot on purpose. Neither is required to record or export a podcast episode — recording, transcription, markers, search and export are unrestricted in the free app.

Common questions

Which quality setting should I use to record a podcast episode?

High is a sensible default for a solo voice — it is well past the point where clarity is limited by the sample rate rather than the room. Ultra makes a larger file without a clearer-sounding voice; standard is a reasonable choice if file size across a season matters more to you than headroom you are unlikely to use.

Can I record an episode longer than an hour without hitting a limit?

Yes. There is no built-in cap on how long a single recording can run or how many recordings you keep — the practical limits are the phone’s storage and battery, not any length restriction in the app.

Will the file I get work in podcast editing software?

Yes. It is a standard .m4a (AAC) file, which opens in essentially any audio editor or timeline. If a specific pipeline needs an uncompressed format, converting from .m4a is a one-step job in most editors.

What happens to a recording if a call comes in during a long solo session?

iOS gives the microphone to the call regardless of what is recording. Voice Studio pauses on the interruption and resumes into the same file once the call ends, so the episode ends up with a gap rather than being cut short or split into two files.

Try it in Voice Studio

Voice Studio records, transcribes on your iPhone, and files each note by time and place — so the thought you had in the car is still findable next month.

Download on the App Store

Free to download · iPhone and iPad · iOS 16.4 or later