Podcast Transcript Format

Paste your raw export, get a clean transcript back in whatever format you actually need.

Podcast Transcript Formatter

A podcast transcript can be formatted as SRT or VTT for video captions, HTML for your website, JSON for the podcast:transcript spec, or plain text for anywhere else — paste or drop your raw export below and the formatter detects which one you have and cleans it up into all six.

Publishing the rest of the episode too?

Write the show notes or episode description with the description builder, or check your cover art against every platform's spec with the cover art checker.

Cutting clips from this episode? Check export settings with the export preset generator and captions against Instagram's Reels safe zone before you post.

What format should I export my podcast transcript in?

It depends where it's going: SRT or VTT for captions on video platforms, HTML for pasting into your website or a host editor that accepts rich text, Markdown for a blog post or docs site, plain text for anywhere formatting would just get stripped anyway, and JSON in the podcast:transcript format if your host or app supports the Podcasting 2.0 namespace directly.

What's the difference between SRT and VTT transcript files?

Both are caption formats with timestamped cues, but VTT is the web standard (used by HTML5 video and most modern platforms) and supports a few extras SRT doesn't, like speaker voice tags. SRT is older and more universally supported by video editors. If a platform accepts both, VTT is usually the safer modern choice.

How should I format speaker labels in a transcript?

"Name:" at the start of each paragraph is the most widely readable convention — bold or ALL CAPS both work too, but plain "Name:" survives being pasted anywhere, including plain-text contexts where bold formatting would just disappear. Keep the same label for the same person throughout; renaming a speaker partway through a transcript is the most common readability mistake.

How long should transcript paragraphs be?

Short enough to scan — a new paragraph on every speaker change and roughly every 100-120 words within one speaker's turn keeps a transcript readable instead of a wall of text. A raw export with a new line every two seconds of audio is the opposite problem: technically accurate, practically unreadable.

Do Apple Podcasts and Spotify auto-generate transcripts for me?

Both platforms can auto-transcribe episodes for in-app search and accessibility, but that auto-generated version usually isn't editable, isn't downloadable in a clean format, and isn't what search engines see on your own website. Uploading your own formatted transcript — via the podcast:transcript tag or directly on your site — is still the only way to control how it reads and get its SEO value.