For podcasters

The episode is done.
The work is not.

Two hours of recording becomes a tight episode, eight clips, captions for all of them, and a set of vertical cuts. All of it described rather than assembled.

  1. 01

    Tighten the conversation

    "Cut the filler and the dead air, but leave the natural pauses between topics."

  2. 02

    Pull the clips

    "Clip the exchange about hiring, and the part where she disagrees with me."

  3. 03

    Make them vertical

    Reframed, captioned, names on screen, sized for wherever they are going.

Unscripted conversation is where filler lives

Podcast recordings carry more filler than any other format, and for a good reason: nobody is reading. Two people thinking out loud produce hesitation, false starts, tangents and long gaps, and that is what makes the conversation feel real.

Some of it should stay. Strip every pause from a conversation and it stops sounding like two people talking and starts sounding like a press release being read at speed. The pause before someone answers a hard question is doing work.

So the instruction has to be able to distinguish, and in plain language it can: "cut the ums and the dead air, but leave the pauses where someone is thinking." That is a distinction no threshold setting can express, and on a two-hour recording it is the difference between an episode that breathes and one that gasps. More on the pacing pass.

Clips are the discovery channel

Almost nobody discovers a podcast by finding a two-hour episode. They see a ninety-second clip, and if it lands they go looking for the source.

Which makes clipping the highest-leverage post-production work there is, and the work most likely to get skipped, because it happens after the episode is already finished and you are already tired.

Ask for them by content: "clip the exchange about hiring", "pull the bit where she disagrees with me", "find the three funniest moments." You choose based on what you know about your audience rather than what a virality score guesses. Each clip is a project you can keep editing, so the one you care most about can get real attention rather than the same automated treatment as the rest.

Two speakers, two sets of problems

Multi-speaker recordings bring specific chores that single-camera video does not.

  • Who is talking. On an interview clip watched muted, a viewer cannot tell who is saying what. Ask for a name lower-third when each person first speaks.
  • Uneven audio. One person is closer to the mic, or louder, or in a worse room. "Even out the levels between us" handles the common case.
  • Different treatment per person. "Tighten me but leave the guest alone" is a real and frequent instruction: you can cut your own rambling more aggressively than your guest's. Give each person's sections a listen afterwards.
  • Speaker framing. For a single wide shot going vertical, deciding who is in frame, which is a case where circling the person you want is faster than describing them.

The video version of an audio show

Most podcasts are now filmed, whether or not anyone planned it that way. Even audio-first shows publish a video version, because that is what YouTube and the social feeds want.

That turns an audio production workflow into a video one, and it is a real change in workload: the tools are heavier, the exports are longer, and the clip work multiplies because each clip now needs framing and captions as well as a trim.

The practical answer is to keep it all in one place. Tighten the episode, pull the clips, reframe and caption them, and export, without moving between an audio editor, a video editor and a clipping tool, each of which knows nothing about the others.

Questions

Can it handle a two-hour recording?
Yes. Long unscripted recordings are the case where the pacing pass saves the most time.
Can viewers tell who is speaking?
Ask for a name lower-third the first time each person talks. It matters most on interview clips, which are usually watched with the sound off.
Can I edit one speaker and not the other?
Yes, ask for it ("tighten the host, leave the guest alone"). Give each person's sections a quick listen afterwards to check the split landed where you wanted.
What about audio-only episodes?
Blob is built around video. For an audio-first show, the fit is the video version and the clips rather than the audio master.
Can I get a transcript?
Yes, ask for one, along with chapter markers based on where the conversation turns.

Cut your next episode

Upload the recording, describe the episode you want, and get the clips too. Free to start.

Start free