Turn podcast audio into
show notes in 60 seconds.

audien·to is an AI audio tool that turns a podcast episode into a complete set of show notes — summary, chapters, pull quotes, and resource links — in about 60 seconds. Unlike generic transcribers that hand you a wall of text, it shapes the episode into something publishable. Free tier, 67 languages, no signup, audio auto-deleted in 72 hours.

Drop your episode in. We’ll write the summary, build chapters, pick the best pull quotes, and list every link mentioned — and hand you the raw transcript too. No signup.

● Record
Drop audio here
or click to choose a file · up to 2h · auto-deleted in 72h
LANGUAGE
required · 67 supported
What you'll get
  • Summary
  • Chapters
  • Quotes
  • Resources
Speakers auto-tagged. Tweak any section after upload from the Options panel — no setup needed up front.
The guide

Why podcast show notes matter

Show notes are the written footprint your podcast leaves on the internet — searchable, shareable, and the first thing listeners see before they press play.

Beyond SEO, they’re a courtesy: a well-structured set of chapters, quotes, and references gives returning listeners a way back into the episode without scrubbing through 40 minutes of audio.

What good show notes contain

  • Summarya single paragraph capturing the thesis.
  • Chapterstimestamped sections readers can jump to.
  • Pull quotessharable lines for social.
  • Resourcesevery book, tool, link mentioned.

How audien·to structures yours

What lands in the doc

  • Episode summaryone tight paragraph, ≤80 words, leading with the question the episode answers.
  • Chapters4–8 timestamped sections. Titles are claims (“Why room beats mic”) not categories (“Gear talk”).
  • Pull quotes2–6 standalone lines, attributed to the speaker, readable without the audio.
  • Resourcesbooks, tools, URLs, and people mentioned — each with its first-mention timestamp.
  • Guests & creditsguest name + role when detected from the transcript or filename.

Writing show notes well

  • Lead with the questionopen with what the episode answers, not the host’s bio. Listeners came for the answer.
  • Chapter titles are claims“The $40 mic test” beats “Microphone discussion.” A reader should know the takeaway from the title.
  • Quotes stand aloneeach pull quote should make sense without the surrounding minute of audio.
  • Always link the resourcesif you mention a book or tool, link it. People will search for it anyway.
  • Skim before playsummary under 80 words, chapters under 10, quotes under 6 — listeners scan first.

Knobs in the Options panel

  • Tonematch the hosts (default) · neutral journalist · concise editor.
  • Chapter granularitybroad (4–6 chapters) · standard (6–10) · fine-grained (every topic shift).
  • Quote count2 / 4 / 8 — fewer = more shareable, more = better SEO surface.
  • Resource groupingsingle timestamped list · grouped by type (books / tools / links / people).
  • Speaker attributionby name (Ada said…) · by role (the host / the guest) when names aren’t reliable.
Why this works

Why modern AI hears what older tools missed

When you upload audio, two AIs go to work. The first one listens. It learned from millions of hours of real speech — accented, overlapping, full of “ums” and brand names that didn’t exist five years ago — so it can hear “Klaviyo,” “Substack,” or “the GPT pipeline” without flinching. The words older tools used to silently mangle come back right.

It hears in context. Instead of guessing one sound at a time, it takes in the whole sentence and uses everything around a tricky word to figure out what was actually said. That’s how a brand-new product name still lands correctly: the words around it tell the AI what kind of sentence it’s in.

And it cleans as it goes. Disfluency — “uh, like, I think… yeah” — doesn’t drop the rest of the sentence on the floor. Punctuation and capitalization come built in, so what you read is prose, not a wall of lowercase. By the time it hands off, the transcript already looks like what a careful typist would have given you.

What happens once we have your words

A raw transcript is the floor, not the ceiling. The second AI reads the whole document the way a careful editor would. It groups related discussion into chapters even when nobody says “moving on.” It surfaces the quote you’d actually screenshot — not the longest sentence on the page. It separates a decision from a tangent, an action item from a passing wish.

That’s the jump older tools couldn’t make: they gave you words, we give you shape. A meeting becomes minutes with owners, dates, and resolved questions. A podcast becomes show notes whose chapters track the real narrative, not the nearest five-minute mark. A voice memo becomes a send-ready email in your voice — not a list of fragments to stitch back together.

Each tool on this page is one of those pairings — the same listening AI up front, the same writing AI behind it, shaped for one specific output. You don’t pick. You don’t tweak. A thirty-second upload comes back as the thing you actually wanted, ready to use.