smithery/hmbown

voiceover-studio

Design custom voices from text prompts and produce professional narration with voice design previews.

Installation

$ npx skills add smithery/hmbown --skill voiceover-studio

Similar popular skills

Related neighbors and high-traction skills in the same topics — useful to compare before installing.

Also in this package

Other skills from smithery/hmbown · top by installs.

npx skills add smithery/hmbown

Browse all from smithery/hmbown

More details

Agent compatibility

Declared targets from SKILL.md / docs. Unmarked agents are not listed — the skill may still install via the CLI.

Claude Code Not declared
Cursor Not declared
Codex Not declared
GitHub Copilot Not declared
Windsurf Not declared
Gemini CLI Not declared
Cline Not declared
OpenCode Not declared

Skill metadata

Parsed from SKILL.md frontmatter.

Allowed toolsvoice_design, voice_list, tts, tts_async_create, tts_async_query, retrieve_file, download_file

Package contents

Files included with this skill beyond the listing page.

  • skill md SKILL.md 2,775 B
  • docs SUMMARY.md 125 B

History

  1. First recorded snapshot · 0 installs

SKILL.md

You are running the Voiceover Studio skill.

Goal

  • Design a custom voice from text descriptions, preview it, and produce full professional narration for any project.

Ask for

  • Project type: commercial, documentary, animation, corporate, audiobook, podcast.
  • Voice characteristics (age, gender, accent, tone, personality).
  • Usage context (broadcast, online, telephone, character voice).
  • Script content or source (file upload or text input).
  • Duration estimate (short spot vs. long-form content).
  • Whether to:

- Design a new voice from scratch - Browse existing voices and customize - Clone from provided samples

  • Quality preference (speech-02-hd for premium, speech-02-turbo for speed).

Workflow

  1. Determine voice approach:

- If designing new: call voicedesign with descriptive prompt (e.g., "warm中年 male voice, slight Southern accent, trustworthy and friendly"). - If browsing: call voicelist to show options with characteristics. - If cloning: request audio sample and call voice_clone.

  1. Generate preview samples:

- Call tts with sample text (2-3 sentences covering different emotions). - Offer 2-3 voice variations for comparison. - Get user feedback and iterate on voice design if needed.

  1. Finalize voice selection:

- Confirm voice_id to use for full production. - Note any specific direction for delivery (energetic, whisper, authoritative).

  1. Process full script:

- If short (<5min): call tts directly with full script. - If long: call ttsasynccreate with script or uploaded file. - Poll with ttsasyncquery until complete. - Download with retrievefile or downloadfile.

  1. Optional: Generate alternate versions:

- Different takes or emotional deliveries. - "Radio edit" (shorter, punchier version) for advertising.

  1. Return production package:

- Voice design specifications (for future consistency) - Preview audio files - Final narration audio - Alternate takes if generated - Timing/word count notes

Response style

  • Be methodical about voice selection—provide samples for comparison.
  • Track voice IDs and specifications for recurring projects.
  • Provide timing estimates based on word count.

Notes

  • Voice design is unique to MiniMax—emphasize this capability.
  • Always get approval on preview before full production.
  • For long-form content, suggest checking pacing mid-way.
  • Offer to generate "sting" or "logo" audio (short signature phrase).
  • Save voice specifications for brand consistency across projects.