ai-podcast-creation

Generate podcasts by synthesizing text-to-speech audio, composing music, and merging audio elements.

688|95|Updated Jan 31, 2026
One-click install
npx skills add https://github.com/inference-sh/skills --skill ai-podcast-creation-inference-sh
Or copy as Structured Prompt for Agent▼
Please help me install this Agent Skill.
Skill: ai-podcast-creation
Source: https://github.com/inference-sh/skills/tree/main/guides/content/ai-podcast-creation
Command: npx skills add https://github.com/inference-sh/skills --skill ai-podcast-creation-inference-sh

SYSTEM DOCUMENTATION & REQUIREMENTS

💡 This Skill includes scripts (resource) and references (resource) components.

What problem does it solve?

This Skill automates the creation of AI-powered podcasts, simplifying the process of generating audio content with text-to-speech, music, and editing.

Core Features & Use Cases

  • Text-to-Speech: Generate speech from text using various AI voices (Kokoro TTS, DIA TTS, Chatterbox).
  • AI Music Generation: Create custom intro/outro music and background tracks.
  • Audio Merging & Editing: Combine voice segments, music, and sound effects into a cohesive podcast episode.
  • Use Case: Produce professional-sounding podcasts, audiobooks, or audio newsletters efficiently, even with multiple AI-generated voices for conversations.

Quick Start

Use the ai-podcast-creation skill to generate a podcast segment with the prompt "Welcome to the AI Frontiers podcast. Today we explore the latest developments in generative AI." using the am_michael voice.

Frequently Asked Questions about ai-podcast-creation

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I generate a podcast using text-to-speech and AI voices?▼

AI podcast creation generates audio by synthesizing text-to-speech with voices like Kokoro TTS, DIA TTS, or Chatterbox. You provide the text script, and the system produces voice segments for your podcast automatically.

Can I create multi-voice conversations for an AI-generated podcast?▼

Yes, AI podcast creation supports multi-voice conversations. You can assign different AI voices to various text segments, allowing automated dialogue generation for interviews or panel-style podcast episodes.

How does AI music generation work for podcast intro and outro tracks?▼

AI music generation creates custom intro, outro, and background tracks for podcasts. The system composes original audio segments that are then merged with voice segments to produce a complete episode.

What is the best way to merge voice segments and background music into a full podcast episode?▼

The best way to merge audio elements is using automated audio merging and editing. The system combines generated voice segments, AI music, and sound effects into a cohesive podcast episode ready for distribution.

Can I use this automated podcast creation workflow for audiobook narration?▼

Yes, automated podcast creation supports audiobook narration and audio newsletter generation. The text-to-speech and audio merging workflows apply directly to producing long-form spoken audio from text scripts.

Do I need any external audio editing software to produce an AI podcast?▼

No, you do not need external audio editing software. The AI podcast creation Skill handles text-to-speech synthesis, AI music generation, and audio merging internally to output a finished podcast episode.