h3-prompt-writing
Write MiniMax H3 video generation prompts for T2VA, I2VA, FL2VA, L2VA, and Ref2VA. Use when rewriting multimodal requests into H3 prompt structures, composing integrated_multimodal_description, overall_soundscape, and non_diegetic_music, aligning keyframes, or defining reference labels for images, videos, and audio.
DeepseekModel
Curated skill
Quality Excellent · 90
v1.0.0
Get
https://deepseekmodel.com/api/download.php?id=minimax-ai-minimax-h3-agents-skills-h3-prompt-writing-skill-md&format=skill
Download .skill
Standard format with system_prompt and model_config, ready for any agent framework
The actual content of the system_prompt field in the .skill file.
name h3-prompt-writing description Write MiniMax H3 video generation prompts for T2VA, I2VA, FL2VA, L2VA, and Ref2VA. Use when rewriting multimodal requests into H3 prompt structures, composing integrated_multimodal_description, overall_soundscape, and non_diegetic_music, aligning keyframes, or defining reference labels for images, videos, and audio. compatibility Portable to any agent that can read local files — no external API calls, MiniMax Hub tools, or proprietary runtime required. The agents/openai.yaml file only adds optional ChatGPT/Codex UI metadata; it does not restrict the skill to OpenAI agents. H3 Prompt Writing Workflow Identify the input mode: T2VA, I2VA, FL2VA, L2VA, or full-reference Ref2VA. For base text/keyframe modes, read references/base-en.txt and follow its final prompt structure. For full-reference mode, read references/ref-en.txt and follow its six-section rewrite format. Preserve the exact field names, section order, labels, and timing notation from the selected guide. Base Modes T2VA: build the full audiovisual timeline from text. I2VA: start from the first frame and develop forward from it. FL2VA: describe the continuous path between the first and last frames. L2VA: infer a plausible opening and converge to the supplied last frame. Use integrated_multimodal_description , overall_soundscape , and non_diegetic_music in the order shown in references/base-en.txt . Full-Reference Mode Ref2VA rewrites use subject_definitions , summary , retention_analysis , detailed_description , overall_soundscape , and non_diegetic_music in that order. Reference labels stay consistent across all sections. Read references/ref-en.txt for label rules, retention analysis, and complete examples. Output Rules Write rewrite sections in English; preserve dialogue, lyrics, and visible scene text in their original language. Describe each shot by composition, subjects, environment, actions, camera, sound, and the exact point where referenced content appears. Avoid plot summaries, unresolved reference labels, and timing that does not match the requested duration.
Keywords that activate this skill. Click one to copy it.
This skill does not provide trigger words.
The downloaded .skill package contains the following fields.
| Field | Description |
|---|---|
| format | Format tag (skill/v1) |
| skill_id | Unique skill ID |
| name | Skill name |
| version | Version |
| description | Description |
| category | Categories (array) |
| trigger_words | Trigger words |
| tags | Tags |
| source | Source |
| source_url | Source URL (this page) |
| exported_at | Exported at (set per download) |
| system_prompt | System prompt body |
| model_config | Model config: provider / model / temperature / max_tokens / top_p |
| examples | Examples |
| install_guide | Import guide for Coze / Dify / Claude / custom frameworks |
The same skill can be exported in different platform formats.