# How to launch a podcast with AI and no studio: from idea to first episode

> Step-by-step guide to launching a podcast with no studio: research with Kagi, voices with ElevenLabs, theme music with Suno, editing and publishing.

- Canonical: https://serchai.com/en/guides/ai-podcast-no-studio/
- Site: Serchai (https://serchai.com) — AI tools comparator
- Language: en
- Updated: 2026-07-24

---

## Tools you will use

- [ElevenLabs](https://serchai.com/en/reviews/elevenlabs/) — Near-human synthetic voices for narration and dubbing.
- [Suno](https://serchai.com/en/reviews/suno/) — Full songs with vocals from a text description.
- [Kagi](https://serchai.com/en/reviews/kagi/) — The paid search engine with no ads and no tracking.

## The steps, in short

1. **Research and script the episodes with Kagi** — Map the topic with an ad-free search engine and turn that research into a script written for the ear.
2. **Generate the voices with ElevenLabs** — Paste the script and produce the narration with a stock voice or your own cloned one, the same voice in every episode.
3. **Compose the theme music with Suno** — Generate the intro music and transitions from a text description and lock them in as the show's sonic identity.
4. **Edit and publish** — Assemble voice and music in a basic editor, listen to the full episode and upload it to a hosting platform with RSS.

> **TLDR:** You can produce a complete podcast with no mic and no studio by chaining three AI tools: Kagi for research and scripting, ElevenLabs for the voices and Suno for the theme music. Final assembly happens in any free audio editor. From first script to published episode fits in one weekend.

This guide is for anyone who wants to launch a podcast but has no decent mic, no quiet room and no budget for voice actors.

By the end you will have a repeatable workflow: research, script, narrate, score and publish. AI covers the production, but the editorial judgment stays with you.

## 1. Research and script the episodes with Kagi

A podcast lives on what it says, not on how it sounds. Before touching any synthetic voice you need the topic mapped: what has already been said, what data exists and which angle is still open. [Kagi](/en/reviews/kagi/) is a paid search engine with no ads and no tracking, and that cleanliness matters when you research: primary sources surface before content-farm rewrites.

Two features fit this job. Lenses filter results by source type, handy for separating press from forums or technical docs. And you can boost or bury specific domains, so trusted sources always rank on top and copycat sites disappear. It starts at 5 dollars a month with a free trial, and its weak spot is exactly that: no permanent free plan. If the price puts you off, our review compares the alternative [Oso.ai](/en/reviews/oso-ai/).

With the research on the table comes the script, and one rule shapes everything downstream: write for the ear, not for the page. Short sentences, figures spelled out as words ("forty-two", not "42") and every proper noun double-checked, because a synthetic voice will read exactly what you type. The next step needs a full verbatim script, not an outline.

## 2. Generate the voices with ElevenLabs

With the script locked, [ElevenLabs](/en/reviews/elevenlabs/) turns it into narration: paste the text and it generates audio with intonation that is hard to tell apart from a human voice. It runs in the browser, freemium, starting at 5 dollars a month. Two paths: a voice from its catalog, ready in minutes, or cloning your own, the better option if the podcast carries your name, though cloning requires verification and a paid plan.

The decision that shapes the result most is consistency. Pick one voice per show (and one per character if you run a two-host format) and never change it between episodes: sonic identity is built by repetition, same as the theme music.

Two practical warnings. Free-plan characters run out fast, so the free tier is fine for auditioning voices but not for producing a full episode. And the commercial license comes with the paid plans: if you plan to monetize the podcast, you need one. Always listen to the complete narration before assembly. The slips cluster around proper nouns, figures and acronyms.

## 3. Compose the theme music with Suno

Every recognizable podcast has an intro, and until recently that meant buying library music or commissioning it. [Suno](/en/reviews/suno/) generates complete songs, vocals and production included, from a text description: you write the genre, tempo and mood you want and within a minute you get a candidate. For a theme tune, ask for it to be instrumental right in the prompt, and if you want a sung tagline in the intro, it takes your own lyrics too.

The practical flow is generating several variants, keeping one and using it forever. The same intro on every episode, plus a trimmed version for transitions, is what makes a show sound like a show. Keep the tool's known limit in mind: control over musical structure is limited, so do not try to direct it second by second. Generate, discard and select.

Suno offers free daily credits and paid plans from 10 dollars a month, on web and iOS. The licensing rule matches the voices one: commercial use requires a paid plan, and a podcast with sponsorship or memberships is already commercial use. In our ranking of [AI audio tools](/en/best-ai/audio/) we compare both tools against the rest of the category.

## 4. Edit and publish

Assembly needs no studio and no paid software: any basic editor like Audacity covers what a podcast requires. Three operations handle 90 percent of the work: aligning the voice with the intro and outro music, ducking the music several decibels under the voice when they overlap, and normalizing the final loudness so no listener has to reach for their volume.

Before publishing, listen to the whole episode with headphones: it is your last chance to catch a mispronounced name or a harsh cut.

To publish you need a podcast hosting platform that generates the RSS feed, which is what Spotify, Apple Podcasts and every other app reads. On cadence, one criterion: an episode every two weeks that actually ships beats a weekly schedule abandoned in month three. If you later turn the same material into training, the workflow in our guide to [launching an online course with AI](/en/guides/online-course-ai/) starts from a similar base.

## Common mistakes

The first one is publishing without listening to the full narration. Synthetic voices rarely fail, but they always fail in the same places: proper nouns, figures and acronyms. Ten minutes of listening saves a public correction.

The second is changing the voice or the theme music between episodes. Every change resets the listener's sonic memory, and without sonic memory you do not have a show, you have loose audio files.

The third is legal: using free-plan voices or music in a monetized podcast. The commercial licenses of ElevenLabs and Suno come with paid plans, and skipping that is a problem that shows up late and ugly.

The fourth is a writing problem: scripting paragraphs meant for the page. What works on screen, long sentences with nested clauses, sounds tangled out loud. If you cannot say a sentence in one breath, cut it.

## Frequently asked questions

### How much does it cost to produce a podcast with this workflow?

The three tools add up to 20 dollars a month: 5 for Kagi, 5 for ElevenLabs and 10 for Suno. All three offer a free trial or starter credits to validate the workflow before paying.

### Can I monetize a podcast with AI voices and music?

Yes, but you need the paid plans of ElevenLabs and Suno, which are the ones that include a commercial license. Free tiers are limited to personal use and do not cover a podcast with ads or memberships.

### Will listeners notice the voice is synthetic?

With a good script, current quality is hard to tell apart from a human voice. What gives an AI podcast away is not the voice, it is a script written for the page: long sentences, unrounded figures and paragraphs with no spoken rhythm.

### Do I need audio-editing skills?

No. Basic podcast assembly means aligning tracks, ducking the music under the voice and normalizing loudness: three operations any free editor does with a ten-minute tutorial. The hard part, researching and writing well, you already did in step 1.
