Skip to main content
ZICQ

Skills ZICQ category:Media h3-prompt-writing

H3 Prompt Writing

Write MiniMax H3 video generation prompts for T2VA, I2VA, FL2VA, L2VA, and Ref2VA. Use when rewriting multimodal requests into H3 prompt structures, composing integrated_multimodal_description, overall_soundscape, and non_diegetic_music, aligning keyframes, or defining reference labels for images, videos, and audio.

8881 installs

Official URL:skills.sh

What this skill does

Intro in this page language first. The official description stays in its original wording; we do not rewrite SKILL.md.

What it does

Write MiniMax H3 video generation prompts for T2VA, I2VA, FL2VA, L2VA, and Ref2VA

When to use it

rewriting multimodal requests into H3 prompt structures, composing integrated_multimodal_description, overall_soundscape, and non_diegetic_music, aligning keyframes, or defining reference labels for images, videos, and audio

How agents load it

Per Agent Skills progressive disclosure: name and description load at startup (~100 tokens); the full SKILL.md body loads when the skill activates; scripts/, references/, and assets/ load only as needed. This file's sections: H3 Prompt Writing; Workflow; Base Modes; Full-Reference Mode; Output Rules; Tips for Better Results. It includes spec-recommended sections: step-by-step instructions.

File analysis

File analysis: besides SKILL.md, the body references references/base-en.txt, references/ref-en.txt. Those resources load on demand.

H3 Prompt WritingWorkflowBase ModesFull-Reference ModeOutput RulesTips for Better Results

Compatibility:Portable to any agent that can read local files — no external API calls, MiniMax Hub tools, or proprietary runtime required. The agents/openai.yaml file only adds optional ChatGPT/Codex UI metadata; it does not restrict the skill to OpenAI agents.

Source category:skills.sh agent-skill

SKILL.md & Agent activation

Official spec ↗
name
h3-prompt-writing
description
Write MiniMax H3 video generation prompts for T2VA, I2VA, FL2VA, L2VA, and Ref2VA. Use when rewriting multimodal requests into H3 prompt structures, composing integrated_multimodal_description, overall_soundscape, and non_diegetic_music, aligning keyframes, or defining reference labels for images, videos, and audio.
compatibility
Portable to any agent that can read local files — no external API calls, MiniMax Hub tools, or proprietary runtime required. The agents/openai.yaml file only adds optional ChatGPT/Codex UI metadata; it does not restrict the skill to OpenAI agents.
  1. DiscoverThe client exposes names and descriptions to the agent.
  2. ActivateYour request or the task context selects the skill and loads its instructions.
  3. Load resourcesReferenced scripts, documentation and assets are used when needed.
Files referenced by the instructions · 2
  • references/base-en.txt
  • references/ref-en.txt

These paths are extracted from the text. Check the upstream package to verify the files exist.

Invocation syntax and available tools depend on your Agent client. Client integration guide ↗

Install this skill

Skills CLI ↗

Choose the target agent and installation scope, keep referenced package files, then verify the skill appears in the client's catalog.

This skill references supporting files. Retrieve the complete directory from the source; copying SKILL.md alone may leave missing dependencies.

Ask your Agent to install

Copy these instructions to a compatible agent and confirm the target directory matches your client.

Install the agent skill "h3-prompt-writing" into my project. The full SKILL.md and official description are at https://zicq.com/en/skills/skl-5239027479db12d7-H3-Prompt-Writing.html
Save it as .cursor/skills/h3-prompt-writing/SKILL.md or .claude/skills/h3-prompt-writing/SKILL.md and keep the frontmatter name and description exactly as-is.
This skill also ships scripts/, references/, or assets/ — fetch the whole folder from https://github.com/minimax-ai/minimax-h3 instead of creating only a SKILL.md.

Full package on GitHub ↗

Install from the terminal · Skills CLI

Requires Node.js and npx. First inspect the repository's skill list to confirm the name.

npx skills add 'https://github.com/minimax-ai/minimax-h3' --list

npx skills add 'https://github.com/minimax-ai/minimax-h3' --skill 'h3-prompt-writing'

The CLI lets you choose the agent interactively. The default scope is the project; use -g for user scope. Confirm package availability with the discovery command, then use npx skills list to inspect installed skills.

Readable layout
--- name: h3-prompt-writing description: Write MiniMax H3 video generation prompts for T2VA, I2VA, FL2VA, L2VA, and Ref2VA. Use when rewriting multimodal requests into H3 prompt structures, composing integrated_multimodal_description, overall_soundscape, and non_diegetic_music, aligning keyframes, or defining reference labels for images, videos, and audio. compatibility: Portable to any agent that can read local files — no external API calls, MiniMax Hub tools, or proprietary runtime required. The agents/openai.yaml file only adds optional ChatGPT/Codex UI metadata; it does not restrict the skill to OpenAI agents. --- # H3 Prompt Writing ## Workflow 1. Identify the input mode: T2VA, I2VA, FL2VA, L2VA, or full-reference Ref2VA. 2. For base text/keyframe modes, read `references/base-en.txt` and follow its final prompt structure. 3. For full-reference mode, read `references/ref-en.txt` and follow its six-section rewrite format. 4. Preserve the exact field names, section order, labels, and timing notation from the selected guide. ## Base Modes - T2VA: build the full audiovisual timeline from text. - I2VA: start from the first frame and develop forward from it. - FL2VA: describe the continuous path between the first and last frames. - L2VA: infer a plausible opening and converge to the supplied last frame. Use `integrated_multimodal_description`, `overall_soundscape`, and `non_diegetic_music` in the order shown in `references/base-en.txt`. ## Full-Reference Mode Ref2VA rewrites use `subject_definitions`, `summary`, `retention_analysis`, `detailed_description`, `overall_soundscape`, and `non_diegetic_music` in that order. Reference labels stay consistent across all sections. Read `references/ref-en.txt` for label rules, retention analysis, and complete examples. ## Output Rules - Write rewrite sections in English; preserve dialogue, lyrics, and visible scene text in their original language. - Describe each shot by composition, subjects, environment, actions, camera, sound, and the exact point where referenced content appears. - Avoid plot summaries, unresolved reference labels, and timing that does not match the requested duration. ## Tips for Better Results - Always match the total duration of the description to the requested video length (4–15 seconds). - Keep reference labels consistent (e.g. ``, ``, ``) across every section. - Prefer concrete visual and audio details over abstract words like "cinematic" or "beautiful". - When using keyframes (I2VA / FL2VA / L2VA), clearly state how the first and/or last frame connects to the timeline.

Related skills

Media

Nano Banana Pro

Generate/edit images with Nano Banana Pro (Gemini 3 Pro Image). Use for image create/modify requests incl. edits. Supports text-to-image + i…

Media

Youtube Watcher

Fetch and read transcripts from YouTube videos. Use when you need to summarize a video, answer questions about its content, or extract infor…

Media

Atxp

Access ATXP paid API tools for web search, AI image generation, music creation, video generation, X/Twitter search, email, and agent account…

Media

Youtube

YouTube Data API integration with managed OAuth. Search videos, manage playlists, access channel data, and interact with comments. Use this …