smithery/hmbown

voice-podcast-kit

Create multi-host podcast episodes with cloned voices, intro music, and transitions.

Installation

$ npx skills add smithery/hmbown --skill voice-podcast-kit

Similar popular skills

Related neighbors and high-traction skills in the same topics — useful to compare before installing.

Also in this package

Other skills from smithery/hmbown · top by installs.

npx skills add smithery/hmbown

Browse all from smithery/hmbown

More details

Agent compatibility

Declared targets from SKILL.md / docs. Unmarked agents are not listed — the skill may still install via the CLI.

Claude Code Not declared
Cursor Not declared
Codex Not declared
GitHub Copilot Not declared
Windsurf Not declared
Gemini CLI Not declared
Cline Not declared
OpenCode Not declared

Skill metadata

Parsed from SKILL.md frontmatter.

Allowed toolsvoice_clone, voice_list, tts, generate_music, list_dir, upload_file

Package contents

Files included with this skill beyond the listing page.

  • skill md SKILL.md 2,078 B
  • docs SUMMARY.md 109 B

History

  1. First recorded snapshot · 0 installs

SKILL.md

You are running the Voice Podcast Kit skill.

Goal

  • Produce a complete podcast episode with distinct cloned voices for each host/guest, intro/outro music, and smooth transitions between segments.

Ask for

  • Episode topic and title.
  • List of speakers/characters (names and optional voice sample descriptions).
  • Whether you should clone voices from provided audio samples, or use existing voice IDs.
  • Episode length target and number of segments (intro, main discussion, listener Q&A, outro).
  • Any music preferences (genre, mood, tempo).

Workflow

  1. Confirm speaker lineup and collect voice samples if cloning is requested:

- If audio files are provided, call voiceclone for each speaker. - If no samples, call voicelist to show available presets and let user choose.

  1. Draft a segment-by-segment outline with speaker assignments and timing.
  2. Generate intro/outro music:

- Call generate_music with appropriate mood (upbeat for intro, winding down for outro).

  1. Write scripts for each segment with clear speaker labels.
  2. For each spoken segment:

- Call tts with the correct voiceid for each speaker. - Use outputformat "mp3" for smooth editing.

  1. Optionally add transition sounds or music beds between segments.
  2. Return a production清单:

- Music files (intro/outro) - Each spoken segment as separate audio files - Full episode script text - Assembly suggestions (timeline order)

Response style

  • Keep track of which voice_id maps to which speaker name clearly.
  • Provide file paths organized by segment and type.
  • Suggest editing workflow but leave final assembly to user.

Notes

  • Voice consistency across episodes is a key value proposition—encourage users to save voice IDs for recurring hosts.
  • If the user wants a sample before full production, offer to generate just the intro + first 2 minutes as a preview.