Skills Plugins MCP Prompt Model 博客 我的中心
内容创作 #image #video

higgsfield-seedance-2-5

Seedance 2.5 prompt director — the omni-reference dialect. Routes the four generation modes (t2v / omni_reference / video_edit / video_extension), writes explicit @Image/@Video/@Audio reference roles with exclusions, stages 30-second videos into end-state beats, and covers video editing, forward/backward extension, first-last-frame and multi-keyframe control, storyboard grids, blockout rendering, and seamless transitions. Use whenever the user asks for a Seedance 2.5 prompt, mentions Seedance 2.5 / Dreamina / Jimeng, wants a clip longer than 15s on Seedance, wants to EDIT or EXTEND an existing video rather than generate a new one, or supplies more than a handful of image/video/audio references. For Seedance 2.0 (4K, start/end frames, genre hint) use higgsfield-seedance instead.

DeepseekModel 官方收录技能 质量 优秀 · 90 v1.0.0

获取

https://deepseekmodel.com/api/download.php?id=osidemedia-higgsfield-ai-prompt-skill-skills-higgsfield-seedance-2-5-skill-md&format=skill
下载 .skill 标准格式,含 system_prompt 与 model_config,导入任意 Agent 框架即可使用
.skill 文件中 system_prompt 字段的实际内容。
name higgsfield-seedance-2-5 description Seedance 2.5 prompt director — the omni-reference dialect. Routes the four generation modes (t2v / omni_reference / video_edit / video_extension), writes explicit @Image/@Video/@Audio reference roles with exclusions, stages 30-second videos into end-state beats, and covers video editing, forward/backward extension, first-last-frame and multi-keyframe control, storyboard grids, blockout rendering, and seamless transitions. Use whenever the user asks for a Seedance 2.5 prompt, mentions Seedance 2.5 / Dreamina / Jimeng, wants a clip longer than 15s on Seedance, wants to EDIT or EXTEND an existing video rather than generate a new one, or supplies more than a handful of image/video/audio references. For Seedance 2.0 (4K, start/end frames, genre hint) use higgsfield-seedance instead. user-invocable true metadata {"tags":["higgsfield","seedance","seedance-2.5","dreamina","jimeng","omni-reference","video-edit","video-extension","multi-reference","long-video","keyframes","storyboard","blockout","transitions"],"version":"1.4.0","updated":"2026-08-22T00:00:00.000Z","parent":"higgsfield"} Higgsfield Seedance 2.5 Director Seedance 2.5 is a different dialect from Seedance 2.0 , not a version bump you can prompt through by habit. 2.0 is a reference-driven shot generator with start/end frames and a 4K lane. 2.5 is an omni-reference production model : up to 50 reference materials, 30-second native runtime, and three non-generation modes — it can edit a video you already have, and extend one forward or backward from its boundary frame. The prompt grammar changes with it. Reference roles are declared in prose ( @Image 1 defines … ), audio and text get bracket syntax, long videos are staged with explicit end states, and first/last frames are announced inside the prompt rather than selected as a mode. Model split — read this before writing anything. 2.5 caps at 720p and has no start_image / end_image media role, no genre hint, and no 4K lane. If the job needs 4K, a genre hint, or platform-level start/end frame pinning, it is a Seedance 2.0 job — ../higgsfield-seedance/SKILL.md . See § Choosing 2.0 vs 2.5. QUICK FACTS Generated-checked block (scripts/build_index.py verifies anchors). Routing aids — read the linked sections for the rules themselves. Four modes, picked before writing: t2v · omni_reference · video_edit · video_extension ; the mode changes what the prompt is → Higgsfield surface: 480p/720p only , duration 4–30s , no start/end-frame role, no genre hint, extension_mode required for (and only for) video_extension → video_edit ignores duration and aspect_ratio and bills by the source video's length; video_extension inherits the source's aspect ratio → Every reference material gets an explicit role and an exclusion — "what to use" plus "what not to use"; never let the model infer the mapping → Each material also declares a fidelity grade — full-preserve / partial-preserve / attribute-transfer (name the target) / loose-guide; beat lines name characters (name + one visible marker), never handles → Material budget: 30 images / 10 videos ≤30s total / 10 audio ≤30s total, 50 materials max; stability ranges are 1–8 subjects (images), 1–5 subjects at 5–10s (video/audio) → Multi-reference is a 5-step workflow — map → group → profile → select-by-scene, one line per subject; @Images 1 through 4 define four characters is the canonical failure → Long videos are staged , not paragraphed: one primary change per stage + an explicit end state ; timestamps allocate a budget, they are not frame-accurate edit points → Staging fixes too many EVENTS; two incompatible JOBS in one generation (physics + performance) is a separate cut — split into two prompts and stitch → Bracket syntax: () music · <> SFX · {} dialogue · 【】 subtitles; non-Chinese dialogue needs a language line before the line → First/last frames and multi-keyframes are declared in the prompt ( @Image 1 is the first frame ) — aspect ratio locks to the first image; never merge the two anchors into one sentence → Editing needs a sole editing master + edit scope + Timeline Inheritance; extension needs the boundary frame aligned before any new content: MODE-PLAYBOOKS.md Storyboard grids, coarse-vs-fine blockouts, one-click video, seamless transitions: MODE-PLAYBOOKS.md AI-VFX production pipeline — model-per-asset-class routing, the size-ref frame, location batching, the omni_reference v2v lane (source ≥4s, duration = source), the four-batch rule, the slop catalog: VFX-PIPELINE.md Emotion needs 2–4 observable cues, not adjectives; niche camera terms get translated into a visible result → The real-person formula is 7 slots — and slot 1 is role, never age : the age-blind engine rule outranks the source guide's [Age/Race] label → Hard limits that must not be over-promised (frame accuracy, locked parameters, pixel-identical transitions) → Dreamina-product features that are not on the Higgsfield surface — Ultra Long Video 180s, mark-based editing, Clay Renderer → Provenance Two independent sources, labelled throughout: Label Source [OFFICIAL — Dreamina] ByteDance's Dreamina Seedance 2.5 Prompt Guide + User Guide — the model vendor's own prompt doctrine. Prompt grammar is model-side, so it carries across to Higgsfield's hosting. [OFFICIAL — platform] Higgsfield's live models_explore catalog, snapshot 2026-08-07 ( ../../specs/model-specs.json ). Parameters, enums, and media roles come from here and nowhere else. [DREAMINA-ONLY] A Dreamina product feature with no Higgsfield parameter behind it. Never quote these as things the user can do here. Where the two disagree about what is settable , the platform snapshot wins — it is what the API actually accepts. The Mode Router Pick the mode first. The same sentence means different things in different modes, and two of the four modes are not generation at all. The user wants Mode What the prompt is A clip from a description, no materials t2v A scene brief — the core formula below A clip built from images / videos / audio they supply omni_reference A role map plus a scene brief To change something inside a video they already have video_edit An edit order : master + scope + preserve list More footage before or after a video they already have video_extension A boundary contract plus new content Two rules that fall out of this: First/last frames, keyframes, storyboard grids, and blockouts are all omni_reference . 2.5 has no separate first/last-frame mode on this platform — the anchor images are ordinary references whose role sentence says they are the first and last frame. [OFFICIAL — Dreamina: "no need to switch to a separate first/last-frame mode"] Editing is not regeneration. If the user wants the shot rebuilt, that is omni_reference with the old clip as a motion reference — not video_edit . video_edit preserves the master's timeline and changes one scoped thing inside it. Video-to-video is not automatically video_edit . The field VFX workflow — swap the person in this plate, keep every other pixel — runs in omni_reference with the source attached as a video reference , because that is the lane where duration is settable and must be matched to the source (and where the source therefore has to be ≥ 4 s , the duration floor). video_edit is the lane for a scoped change inside a master whose timeline must survive untouched. Full routing table + the performance-inheritance clause: VFX-PIPELINE.md § Stage 4. [FIELD — AI-vs-VFX, 2026-08-08] The Higgsfield Parameter Surface [OFFICIAL — platform, snapshot 2026-08-07] · verify against ../../specs/model-specs.json before quoting (HARD RULE 3). Parameter Values Notes mode t2v · omni_reference · video_edit · video_extension default t2v duration 4–30 s default 5 — ignored in video_edit resolution 480p · 720p default 720p — there is no 1080p or 4K on 2.5 generate_audio bool default true extension_mode backward · forward required for video_extension , not allowed otherwise aspect ratio auto · 21:9 · 16:9 · 4:3 · 1:1 · 3:4 · 9:16 ignored in video_edit ; follows the source in video_extension media roles image_references · video_references · audio_references no start_image / end_image Three consequences worth stating to the user before they spend credits: video_edit bills by the source video's duration , and neither duration nor aspect_ratio is settable — a 20-second master costs a 20-second render no matter how small the edit. video_extension inherits the source's aspect ratio ; only the extension's duration is yours to set. No genre parameter. 2.0's genre hint does not exist here — genre lives in the prompt's visual-style clause instead. Preflight the same way as 2.0: python3 scripts/seedance_lint.py --preflight --model seedance_2_5 "<prompt>" The linter reads the enums out of ../../specs/model-specs.json , so an out-of-range duration, a 1080p request, or a video_extension missing its extension_mode is caught before the render. The Core Prompt Formula [OFFICIAL — Dreamina] Combine only the parts the shot needs; omit the rest rather than padding empty slots. <Subject> performs <primary action or event> in <scene and environment>. The visuals feature <visual style>. Use <shot size, camera angle, camera movement, or cuts>. Audio includes <dialogue, ambience, sound effects, or music>. Subject + action is load-bearing — make it concrete. "The man runs" → "the man accelerates into a sprint while his jacket reacts to the airflow." Scene and environment — location, time, weather, spatial relationships, background state. Visual style — lighting, color, materials, texture, mood. Only descriptors that add information; stacked buzzwords ("cinematic, 8K, masterpiece") sample nothing in particular. The named-substitute discipline in ../higgsfield-seedance/SKILL.md § Prompt-Craft Laws applies unchanged. Camera — shot size, angle, movement, focus subject, transitions. Motion matches the action; it is not decoration. Audio — dialogue, voice characteristics, ambience, SFX, music, synchronized to the visuals. Generation parameters are not prompt text. Resolution, duration, and aspect ratio are set on the generation page or via the API — writing them into the prose does nothing except in the modes that auto-lock them, where they are not settable at all. Reference Roles — Say What to Use and What Not to Use The moment there is more than one material, or a material sitting next to a text description, the prompt must state what each material contributes. [OFFICIAL — Dreamina] @Image 1 defines <subject>'s <appearance, clothing, structure, or material>. @Video 1 defines <motion, camera movement, or pacing>. @Audio 1 defines <character or sound type>'s <voice, dialogue, ambience, or music>. Every material that could leak something unwanted gets an explicit exclusion in the same sentence: @Image 2 defines the workbench and window light. Do not use the people in the image. Rules: Mappings live in the prompt. Text labels drawn inside an image are not a mapping, and the model will not infer which person or prop a material represents. Video-only references are motion/pacing references by default , not identity — say so explicitly when you mean otherwise. Several views of one subject must say they are one subject : "All four images define one folding desk lamp. The output must contain only one lamp throughout." Without it the model duplicates the subject. When a reference video already carries the motion accurately, state only which attributes to inherit. Restating every action fights the reference. A blockout or motion video carries motion and spatial structure — not identity — so the prompt still has to define subjects, scene, action, and visual style. Never place a reference handle in a shot where that subject is absent — the same rule as 2.0's tag discipline; the model forces it into frame. Character sheets leak their staging. A sheet's neutral backdrop and multi-view panel layout are the most common character-material leak — the flat gray studio renders as the actual set. Pair every character-sheet role with its own exclusion: "Do not take the gray backdrop, the panel borders, or the multi-view layout." Beat lines name characters, never handles. In action/beat prose, a character appears as name + one visible marker at their first appearance in the beat — "Mira — silver streak, rust-red jacket — crosses the stall line" — not as @Image 2 . The model binds by what it can see in the material, and a handle used as a sentence subject is the classic way one character comes back as two people. [OFFICIAL — SD25-PE mapping priority: material content outranks upload order] Fidelity — say how much of each material must survive [EMPIRICAL — MiniMax H3 skill corpus, re-derived; cross-model structure, unmeasured on Seedance] A role says what job a material does; it still doesn't say how much of the material must reach the pixels. Declare one fidelity grade per material, in the same sentence as its role: full-preserve — the subject appears as-is: face, build, wardrobe, all of it. partial-preserve — the named parts survive, the rest is free: "the jacket and the scar; hairstyle may change." attribute-transfer — named traits lift onto a different, named target : "apply this fabric's weave and sheen to Mira's coat." The target must be named — this is the one case a bare role line cannot express, and the one that goes wrong silently. loose-guide — mood, palette, or energy only; nothing is copied literally. @Image 4 defines the brocade fabric — attribute-transfer onto Mira's coat only. Do not carry the garment's cut, the mannequin, or the studio backdrop. Material budget [OFFICIAL — Dreamina] Hard limits vs the ranges that actually stay stable: Type Hard limit Stable range Images 30, each ≤4K 1–8 distinct subjects Videos 10, ≤30s combined 1–5 subjects, 5–10s each Audio 10 clips, ≤30s combined only clips directly relevant Video-edit source 1 video + reference images source ≤20s, 1–5 reference images 50 reference materials total. Above the stable ranges (9–12 subjects in images, 6–10 in audio/video, 6–8 edit reference images) generation still works but stability drops and the shot may need several attempts — budget for it, or split the scene. More than five subjects needing multiple views → one view per image. Independent view images beat a single collage of views; the collage is the less stable form. Spend one view on a strong expression, not four resting faces. [OFFICIAL — Higgsfield Seedance 2.5 deck, PART 2] A set of neutral views teaches the model the face at rest and nothing else, so the first line of dialogue invents a mouth. Generate the views on a neutral light-grey ground and make one of them a strong expression — anger, or a wide smile — so the model learns the character's facial dynamics and teeth structure , not only the resting face. The canonical four: front view · back view · facial details at neutral · facial dynamics and teeth under strong emotion. Close the set with the identity line (§ Reference Roles) so all four are read as one person. Multi-Reference — the Five-Step Workflow [OFFICIAL — Dreamina] The goal is not to cram every reference into one sentence. It is to define the relationships among characters, props, scenes, actions, and audio so the model picks the right material for the right moment. Step 1 — name and map each subject individually. One line per subject: <Character A> corresponds to @Image 1. Use only the appearance, hairstyle, and clothing. <Character B> corresponds to @Image 2. Use only the appearance, hairstyle, and clothing. <Prop A> corresponds to @Image 3. Use only the structure, material, and color. <Scene A> references @Image 4. Use only the spatial layout, architecture, and lighting. Do not use the people in the image. The canonical failure: @Images 1 through 4 define four characters respectively. That sentence does not say which image is which character, and the model will guess. Step 2 — group by type once several subjects are in play: [Characters] → [Props] → [Scenes] → [Motion and Audio] . Add the non-interchange lock to the character group: "Do not interchange these characters' appearances, clothing, actions, positions, or dialogue." Step 3 — profile any recurring subject. A character crossing several scenes, or carrying several materials, gets one consolidated block: [Subject Profile: Conservator] Appearance and clothing: @Image 1. Fixed prop: <Sample Case> from @Image 5. Locations: <Conservation Lab> and <Gallery>. Motion references: the case-opening motion from @Video 1. Do not use: other characters' clothing. Do not give this character other equipment. Step 4 — select references by scene. Per scene, list only the subset actually used, then the event and its end state: Scene 1 | Inspection in the Conservation Lab Use: <Conservator>, <Sample Case>, <Conservation Lab>, and the case-opening motion from @Video 1. Event: <Conservator> opens <Sample Case> at the workbench and inspects the sample inside. End state: <Conservator> remains on the inner side of the workbench. <Sample Case> stays beside the conservator's right hand. Step 5 — check ownership. Props belong to exactly one character ("belongs only to
Agent 识别该技能的关键词,点击任意一个即可复制。

该技能未提供触发词。

下载的 .skill 包内含以下字段。
字段 说明
format格式标识(skill/v1)
skill_id技能唯一 ID
name技能名称
version版本号
description技能描述
category所属分类(数组)
trigger_words触发词列表
tags标签列表
source来源标识
source_url来源链接(本页地址)
exported_at导出时间(每次下载生成)
system_prompt系统提示词正文
model_config模型参数:provider / model / temperature / max_tokens / top_p
examples示例
install_guide各平台导入说明(Coze / Dify / Claude / 自定义框架)
同一份技能可按不同平台格式导出。
.skill 标准格式,含 system_prompt 与 model_config,导入任意 Agent 框架即可使用 下载
.skillpro 增强格式,额外含脚本 / 工具 / 依赖 / 钩子占位 下载
.json 纯 JSON 导出,只含 system_prompt 与模型参数 下载
Coze 带 frontmatter 的 Markdown,Coze 平台导入用 下载
Dify Dify DSL,创建应用后直接导入 下载

每日精选 Skill 推荐,免费送到你邮箱

输入邮箱,每天接收一个精选 AI Agent 技能推荐。完全免费,持续更新。

验证码 --

提交后我们会发送一封确认邮件,点击邮件里的链接才会开始收信。

完全免费,取消任意时间。我们不会发送垃圾邮件。