Podcast: the video podcast studio

Upload an episode and get a video podcast. Automatic speaker detection finds up to 4 voices, seats each host on an animated stage, and the visuals move with whoever is talking.

Podcasts are audio, but the platforms that grow them are video. Podcast, the video podcast studio, turns one upload into an episode you can put on YouTube or deliver to Spotify, without opening an editor.

How does speaker detection work?

Upload the episode as it is, mixed, mono or stereo, and tell the studio how many people are talking, up to 4. It listens to the whole recording and works out when each of them starts and stops and where they overlap, from the voices alone. There is no script to paste and no timeline to mark up. 4 seats cover a host, a co-host and guests.

More detail, including what happens when an intro jingle or an advert is mistaken for a person, is on speakers and voice detection.

What does the finished episode look like?

  • Each speaker has a photo, a name and a role
  • Their portrait animates while they talk and settles when they stop
  • An animated scene or your own photo sits behind them
  • Titles, logos and images sit on top, in layers you control

What does it export?

1920 x 1080 at 50 frames a second, with the episode's audio encoded as AAC at 192 kbps, for episodes up to 3 hours.

Last updated 2026-09-15

LimePulse is not affiliated with, sponsored or endorsed by Spotify AB or Apple Inc. Product names are used only to describe what the software makes files for.