Text to Speech for Podcasts

Paste a script, pick a voice, and get an MP3 back. Useful for solo narration, daily briefings, and — with the multi-speaker studio — scripted interviews or dialogue where each speaker needs a different voice.

No login • 580+ voices in 75+ languages • Up to 5,000 characters per generation

What this is actually good for

If your show is a script you read out loud — a daily news briefing, a solo explainer, a research summary — the single-narrator studio handles it directly: paste the script, choose a voice and pacing, generate, download the MP3, drop it into your editor.

If the format has more than one voice — a scripted "interview," a two-host cold open, a dramatized segment — use the multi-speaker studio instead. You write the script as labeled lines, assign a different AI voice to each speaker, and it renders as one combined audio file rather than several clips you have to stitch together yourself.

What this isn't good for: unscripted, conversational banter. AI narration reads what you write — it doesn't riff. If your show's appeal is two people reacting to each other in the moment, that still has to be recorded live.

Controlling pacing with pause tokens

Podcast narration lives and dies on pacing — a script read at one flat speed sounds robotic fast. The editor supports three pause lengths you can drop directly into your text: [[p400]] for a short beat, [[p700]] for a sentence break, and [[p1200]] for a segment transition.

Example: "Today's top story — a shift in interest rates. [[p700]] Here's what it means for you. [[p400]] First," — the two pauses do the work a human host would do naturally with breath and emphasis.

Which mode fits your format

Single voice

  • Daily news or market briefings read from a script
  • Solo explainer or commentary episodes
  • Research summaries and monologue-style shows

Multi-speaker

  • Scripted "interview" segments written as Q&A
  • Two-host cold opens or scripted banter
  • Dramatized or character-driven fiction podcasts

From script to MP3

1. Paste or import the script

Type directly, or import a PDF/Word doc. One generation covers up to 5,000 characters — longer episodes get split into a couple of sections.

2. Pick voice and pacing

Choose a narrator from 580+ voices, adjust pitch/rate/volume, add pause tokens where you want a beat.

3. Generate and download

Preview the MP3 in the browser, then download it straight into your editing timeline.

One honest note

Listeners can usually tell a script is AI-narrated within the first few sentences, especially on longer-form shows. That's not a dealbreaker for briefings, summaries, or explainer content where the value is the information — but if your show's whole appeal is a distinct human voice or unscripted chemistry between hosts, text to speech won't replace that, and pretending otherwise to your audience is a bad idea.

Frequently asked questions

Can I use this for a two-host podcast?

For a scripted version, yes — write it as labeled lines and use the multi-speaker studio to assign a different voice to each host. For live, unscripted back-and-forth, no — that still needs to be recorded.

How long can one episode be?

Each generation covers up to 5,000 characters (roughly 800–1,000 words, depending on sentence length). For a full episode, generate in sections and combine the MP3s in your editor.

What file do I get?

An MP3, ready to drop into whatever you edit and publish with.

Do I need an account?

No. Open the tool, paste your script, generate.

Turn your next script into audio

Open the Studio