Podcast Transcription
Free AI transcript for every episode
Drop your episode from any recording tool. Get a clean, timestamped transcript with AI summary ready for show notes, SEO recap, and YouTube captions.
Drop your file here
or click to browse
MP3 · MP4 · WAV · M4A · OGG · WEBM · FLAC · Max 25MB · Max 30 min
How podcast transcription works
Example podcast transcript
- •Episode 42, focus on independent podcasting in 2026.
- •Guest: 8-year weekly show host.
- •Discussion: distribution changes 2018 vs 2026.
- •Attention economy the new bottleneck.
- Draft show notes from summary
- Extract 3 pull quotes for social
- Publish SRT to YouTube video version
Illustration. Real transcripts show timestamps and text without speaker labels.
Why Mictoo for podcast transcription
Podcast transcription scenarios
Tips for cleaner podcast transcripts
- Bounce a voice-only stem when your DAW allows it.
- A voice-optimised 64 kbps mono MP3 transcribes just as well.
- For interview episodes, remove music beds before transcription.
- Review host and guest names once, they appear consistently after.
Frequently asked questions
What podcast platforms does Mictoo work with?
All of them. We accept any audio or video file regardless of host. Common: Riverside, Zencastr, SquadCast, Anchor, Buzzsprout, Libsyn. Free per file up to 60 MB signed in.
Can I get the transcript as show notes automatically?
The AI summary appears alongside the transcript automatically. Use it as the first draft of your show notes then trim to fit your voice.
My episode is over 60 MB. What now?
Sign in to auto-split longer files, or downsample first with ffmpeg -i episode.mp3 -b:a 64k -ac 1 out.mp3. A 60-minute podcast becomes ~28 MB with no accuracy loss.
Can I download SRT for the YouTube video version?
Yes. SRT and VTT include timestamps aligned to the original audio. Drop into YouTube Studio directly.
Does Mictoo transcribe non-English podcasts?
Yes. Whisper large-v3 supports 50+ languages with auto-detection. For short episodes or unusual accents, pick the language explicitly.
Are episodes stored on your servers?
No. The audio streams to the transcription provider, gets processed once, and is dropped from memory. Transcripts are only stored if you sign in.
Can I identify who is speaking (host vs guest)?
No. The current transcript is continuous text with timestamps and no automatic speaker labels.