minimax-ai/minimax-h3

h3-prompt-writing

Write MiniMax H3 video generation prompts for T2VA, I2VA, FL2VA, L2VA, and Ref2VA.

All-time #2137 Trending #1011 Hot #3593 First seen Aug 5, 2026
8-week activity · all time api

Installation

$ npx skills add minimax-ai/minimax-h3 --skill h3-prompt-writing

Summary

  • Write MiniMax H3 video generation prompts for T2VA, I2VA, FL2VA, L2VA, and Ref2VA.
  • Use when rewriting multimodal requests into H3 prompt structures, composing integrated_multimodal_description, overall_soundscape, and non_diegetic_music, aligning keyframes, or defining reference labels for images, videos, and audio.

Similar popular skills

Related neighbors and high-traction skills in the same topics — useful to compare before installing.

Also in this package

Other skills from minimax-ai/minimax-h3.

npx skills add minimax-ai/minimax-h3

Browse all from minimax-ai/minimax-h3

More details

Agent compatibility

Declared targets from SKILL.md / docs. Unmarked agents are not listed — the skill may still install via the CLI.

Claude Code Not declared
Cursor Not declared
Codex Declared
GitHub Copilot Not declared
Windsurf Not declared
Gemini CLI Not declared
Cline Not declared
OpenCode Not declared

Repository health

Stars 8.3K
License MiniMax H3 Community License Agreement
Default branch main
Open issues 30
Status Active

Skill metadata

Parsed from SKILL.md frontmatter.

CompatibilityPortable to any agent that can read local files — no external API calls, MiniMax Hub tools, or proprietary runtime required. The agents/openai.yaml file only adds optional ChatGPT/Codex UI metadata; it does not restrict the skill to OpenAI agents.
Declared agents codex

Package contents

Files included with this skill beyond the listing page.

  • skill md SKILL.md 2,623 B
  • docs SUMMARY.md 342 B

History

  1. First seen on skills.sh
  2. First recorded snapshot · 7,200 installs

SKILL.md

H3 Prompt Writing

Workflow

  1. Identify the input mode: T2VA, I2VA, FL2VA, L2VA, or full-reference Ref2VA.
  2. For base text/keyframe modes, read references/base-en.txt and follow its final prompt structure.
  3. For full-reference mode, read references/ref-en.txt and follow its six-section rewrite format.
  4. Preserve the exact field names, section order, labels, and timing notation from the selected guide.

Base Modes

  • T2VA: build the full audiovisual timeline from text.
  • I2VA: start from the first frame and develop forward from it.
  • FL2VA: describe the continuous path between the first and last frames.
  • L2VA: infer a plausible opening and converge to the supplied last frame.

Use integratedmultimodaldescription, overallsoundscape, and nondiegetic_music in the order shown in references/base-en.txt.

Full-Reference Mode

Ref2VA rewrites use subjectdefinitions, summary, retentionanalysis, detaileddescription, overallsoundscape, and nondiegeticmusic in that order. Reference labels stay consistent across all sections.

Read references/ref-en.txt for label rules, retention analysis, and complete examples.

Output Rules

  • Write rewrite sections in English; preserve dialogue, lyrics, and visible scene text in their original language.
  • Describe each shot by composition, subjects, environment, actions, camera, sound, and the exact point where referenced content appears.
  • Avoid plot summaries, unresolved reference labels, and timing that does not match the requested duration.

Tips for Better Results

  • Always match the total duration of the description to the requested video length (4–15 seconds).
  • Keep reference labels consistent (e.g. <Picture 1>, <Video 1>, <Audio 1>) across every section.
  • Prefer concrete visual and audio details over abstract words like "cinematic" or "beautiful".
  • When using keyframes (I2VA / FL2VA / L2VA), clearly state how the first and/or last frame connects to the timeline.