Spotify Podcast Transcription: Complete Guide
Spotify is the world's largest podcast platform, but getting a usable text version of an episode is still frustrating. Spotify's in-app transcripts cannot be exported, copied, or searched outside the app — so to get a transcript you can actually work with, you download the episode audio and run it through a transcription service like BrassTranscripts. This guide walks through why Spotify's transcripts are locked, how to obtain the audio legitimately, and how to turn it into a speaker-labeled text file you can search, edit, and repurpose.
Quick Navigation
- Does Spotify Have Built-In Transcripts?
- How to Get the Audio: Downloading Spotify Episodes
- Which Transcription Path Should You Choose?
- Transcription Options for Podcast Audio
- Step-by-Step: Transcribing a Podcast with BrassTranscripts
- Getting Better Podcast Transcripts
- Working with Podcast Transcripts
- Legal Considerations
- Frequently Asked Questions
Does Spotify Have Built-In Transcripts?
Spotify displays auto-generated transcripts for select podcasts, but this feature is designed to keep you listening inside the app — it offers no export, download, or copy function for listeners. Creators can access and download their own episode transcripts as VTT files; everyone else can only view text synced to playback on screen.
According to Spotify's transcripts documentation, transcripts appear only for episodes where the feature is enabled, and availability remains inconsistent across the catalog.
What Spotify transcripts offer listeners:
- Auto-generated text synced to audio playback
- In-app display while listening
- Searchable text within a single episode
What listeners cannot do:
- Download or export the transcript
- Copy or paste the full text
- Search or index transcripts across episodes
- Use the text for research, citation, or content repurposing
The gap: if you are a listener rather than the show's creator, Spotify gives you no way to extract the transcript. For every use case beyond casual in-app reading — research, accessibility, repurposing, archiving — you need to transcribe the audio yourself.
How to Get the Audio: Downloading Spotify Episodes
Because Spotify provides no direct download links, the reliable path is to obtain the same episode from its original public source rather than from Spotify itself. Nearly every podcast is distributed through an RSS feed, which means the same audio file usually lives on the show's own website or another open directory.
Option 1: Original Source Download (Recommended)
Most podcasts publish episodes on multiple platforms. Check these first:
Podcast website
- Search for "[Podcast Name] official website"
- Open the episode archive or "Listen" page
- Many sites offer a direct MP3 download link or a player with a download button
RSS feed
- Find the show's RSS feed URL (often linked in its website footer)
- Open the feed in a browser or feed reader
- Each
<enclosure>entry contains the episode's audio file URL - Right-click the audio link to save the file
Apple Podcasts (desktop)
- Open Apple Podcasts on macOS
- Subscribe to the show and download the episode
- Locate the cached file through your library
Option 2: Alternative Podcast Apps
Several podcast players allow offline downloads: Pocket Casts, Overcast, and Castro all support downloading episodes within the app. Note that these apps store audio in app-managed locations and formats, so you may need to export or locate the cached file before transcribing.
Option 3: Browser-Based Capture
For shows with no direct download option, audio from a web player can sometimes be captured, or you can record the system audio output while the episode plays (lower quality). Verify the legality of any capture method and use it for personal purposes only.
Which Transcription Path Should You Choose?
The right transcription method depends on what you need from the text — quick personal notes, publishable content, courtroom-grade accuracy, or accessibility captions all point to different tools. The decision framework below matches common goals to the most cost-effective path so you do not overpay for accuracy you do not need.
| Your goal | Best path | Why |
|---|---|---|
| Personal notes / studying | AI transcription | Fast and inexpensive; minor errors are acceptable for private use |
| Content repurposing (blog, social, show notes) | AI transcription + light edit | Speaker labels and timestamps speed up repurposing; you polish only what you publish |
| Accessibility captions (SRT/VTT) | AI transcription | Produces caption files directly; review for on-screen accuracy |
| Verbatim legal/medical record | Human or hybrid | When 99%+ verbatim accuracy is legally required, human review is worth the cost |
| Poor audio (phone guest, heavy noise) | Hybrid (AI draft + manual fix) | AI handles the bulk; you correct the hard passages |
For most podcast listeners and creators, AI transcription with automatic speaker identification is the sweet spot: it is inexpensive, returns results in minutes, and separates each speaker without manual labeling.
Transcription Options for Podcast Audio
Once you have the audio file, three transcription paths cover essentially every need, trading speed and cost against maximum accuracy. AI transcription is fastest and cheapest, human transcription is most accurate for difficult audio, and a hybrid of the two balances both.
AI Transcription Services
Best for: speed, cost, and automatic speaker identification.
| Service | Turnaround | Speaker ID | Pricing model | Output formats |
|---|---|---|---|---|
| BrassTranscripts | 1–3 min per hour of audio | Automatic | Flat per file ($2.50 / $6.00) | TXT, SRT, VTT, JSON |
| Otter.ai | Real-time | Automatic | Subscription | TXT, SRT |
| Descript | Minutes | Automatic | Subscription (bundled with editing) | TXT, SRT, and editor project |
For a full cost breakdown across providers, see the AI transcription pricing comparison. Podcast-specific advantages of AI services include automatic speaker separation, robust handling of conversational speech, efficient processing of long episodes, and multiple export formats from a single upload.
Manual Transcription
Best for: perfect verbatim accuracy or audio too difficult for automated tools. Manual transcription takes roughly 4–6 hours of work per hour of audio, or $1–3 per minute for a professional human service, and gives you complete control over formatting and accuracy.
Hybrid Approach
Best for: important content that needs accuracy without full manual cost. Run AI transcription first for a fast, inexpensive draft, then review and correct only the passages that matter — typically cutting editing time to a fraction of full manual transcription while reaching near-verbatim quality.
Step-by-Step: Transcribing a Podcast with BrassTranscripts
Transcribing a downloaded episode is a four-step process that takes only a few minutes of processing per hour of audio. BrassTranscripts requires no subscription — you preview the result before paying and download every format at once.
Step 1: Prepare your file
- Supported formats: 11 common audio and video types, including MP3, M4A, WAV, AAC, FLAC, OGG, Opus, and MP4
- Maximum file size: 450 MB
- No enforced duration limit (a single 450 MB file typically covers several hours of compressed podcast audio)
Step 2: Upload and preview
- Go to brasstranscripts.com
- Upload your podcast episode
- Review the 30-word preview to confirm the audio was read correctly
- Check the number of speakers detected
Step 3: Process and download
- Pay the flat rate — $2.50 for files 1–15 minutes, $6.00 for 16 minutes and up (any length)
- Wait for processing (typically 1–3 minutes per hour of audio)
- Download all four formats: TXT, SRT, VTT, JSON
Step 4: Review the output The transcript includes full text with speaker labels (Speaker 1, Speaker 2, …), segment-level timestamps, and a format for every use — plain text for reading, SRT/VTT for captions, and JSON for programmatic processing.
Getting Better Podcast Transcripts
Two factors drive podcast transcript quality more than anything else: the cleanliness of the source audio and how well the tool separates speakers. Getting both right is the difference between a transcript you can publish and one you have to rewrite.
Audio Quality Comes First
Studio-recorded podcasts with dedicated microphones transcribe cleanly, while remote interviews with a phone-quality guest, loud music beds, or overlapping crosstalk introduce the most errors. If you produce the show, recording each speaker on a separate clean track dramatically improves results. Our audio quality guide covers the recording practices that most improve transcription accuracy.
Getting Speaker Names Right
AI diarization labels voices generically as Speaker 1 and Speaker 2; you then rename them to the actual host and guests in the output. Interview and panel formats with more voices carry a higher chance of speaker confusion, so review the labels on multi-guest episodes. See how to get speaker names in transcripts and the speaker identification guide for the full workflow.
Handling Long Episodes
Multi-hour episodes transcribe fine as a single upload within the 450 MB limit, but if a file exceeds that size, split it with a free tool like Audacity and keep speaker labels consistent across the parts when you recombine.
Working with Podcast Transcripts
A single podcast transcript is a reusable content asset — one episode's text can seed a blog post, a week of social snippets, show notes, and a newsletter. Turning audio into searchable text unlocks repurposing, research, and accessibility uses that the audio alone cannot serve.
Content Repurposing
- Blog posts: extract key topics, edit into article format, add headers and context, and link back to the episode
- Social media: pull compelling quotes into text snippets sized for each platform, with an episode link
- Show notes: summarize discussion points, list mentioned resources, and add timestamps for key moments
- Newsletters: turn episode highlights into a digest that delivers value without requiring a full listen
The podcast content repurposing guide and the content-creator transcription workflow go deeper on turning one recording into many assets.
Research and Study
A searchable transcript library lets you find exactly when a topic was discussed, locate a specific guest appearance, and quote accurately with timestamps for citation. This turns a back catalog of episodes into a personal, indexable knowledge base.
Accessibility
Transcripts and caption files make podcast content usable by deaf and hard-of-hearing audiences, support non-native speakers, and let anyone consume content by reading — which is faster than listening for many people. A text version also preserves the content if the audio ever goes offline.
Legal Considerations
Transcribing a podcast you did not create sits within copyright law, and the line runs between private use and publication. Personal, non-commercial transcription for study, notes, or accessibility generally qualifies as fair use in the United States, while publishing or selling transcript text requires permission from the rights holder.
Personal Use
Generally acceptable without explicit permission: personal study and note-taking, accessibility accommodation, private research, and non-commercial archiving. The U.S. Copyright Office's Fair Use Index documents how courts weigh purpose, amount used, and market effect in these cases.
Commercial Use
Requires permission from the copyright holder: publishing transcript excerpts, creating derivative content for sale, using transcripts in commercial products, or redistributing full transcripts.
Best Practices
- Check terms of use — review the show's stated policies
- Credit sources — always attribute the original creators
- Seek permission — when in doubt, ask
- Understand fair use — know the guidelines in your jurisdiction
- Support creators — back the podcasts you rely on
Frequently Asked Questions
Can I transcribe any Spotify podcast?
Technically yes, if you can obtain the episode audio from a legitimate source. Whether you should depends on use: personal transcription for notes, study, or accessibility generally falls under fair use, while commercial use or republication requires permission from the copyright holder.
Why doesn't Spotify let me download transcripts?
Spotify's transcripts are built to enhance in-app listening, not to enable text extraction. Keeping transcripts view-only inside the app retains users in Spotify's ecosystem and limits redistribution of creators' content, which is why listeners get no export function.
How accurate is AI transcription for podcasts?
Accuracy depends primarily on the source audio rather than the AI model. Professional podcasts recorded with good microphones typically produce clean, usable transcripts, while episodes with phone-quality guests, heavy music, or crosstalk need more correction.
Can AI identify different speakers in a podcast?
Yes. Automatic speaker identification (diarization) distinguishes voices by their acoustic characteristics, and BrassTranscripts labels each speaker automatically as Speaker 1, Speaker 2, and so on. You then rename those labels to the real host and guest names in the output.
What about podcasts in other languages?
AI transcription supports many languages with automatic language detection, and BrassTranscripts handles 99+ languages. The episode is transcribed in the language spoken; translating the resulting text into another language is a separate step.
How do I handle very long podcast episodes?
Long episodes transcribe fine as a single upload as long as the file stays under the 450 MB limit. If a recording exceeds that size, split it into parts with a free tool like Audacity, transcribe each part, and keep speaker labels consistent when you combine them.
Is the transcript ready to publish as-is?
Treat an AI transcript as a fast first draft. It is accurate enough for personal notes and search immediately, but before publishing you should rename speaker labels, fix any terms the model misheard, and lightly format the sections you plan to share.
Related Resources:
- Podcast Content Repurposing Guide
- Content-Creator Transcription Workflow
- How to Get Speaker Names in Transcripts
- Speaker Identification Guide
- Audio Quality Secrets for Perfect Transcription
- AI Transcription Pricing Comparison
Ready to transcribe a podcast? Upload your episode and get a speaker-labeled transcript in minutes. No subscription required.