Releases: OSideMedia/higgsfield-ai-prompt-skill
Release list
v3.22.1
Post-release audit fix pass. A two-agent audit of v3.22.0 (cross-reference integrity + repo hygiene) found no broken references but six substantive contradictions between new and pre-existing rules, plus stale downstream copies of the old 200-word rule. All fixed.
Fixed
- The ~2,500-character "practical limit" scoped as ZH-derived (seedance 1.11.1 § Shot density): it contradicted the same file's new Field-calibration medians (1,433–2,059w). EN block prompts have no character analogue; ZH keeps the 1,800-char hard cap. Section now cross-links the fuller shotlist-director density heuristic.
- Runtime-default contradiction resolved (shotlist-director 1.1.1): auto-enrichment no longer silently defaults to 8s — inside a shotlist the 15s envelope law governs; standalone prompts keep seedance's "always ask for runtime, never default."
- Duration ladder vs 15s target reconciled (shotlist-director): 15s is the default envelope, not a straitjacket — a scene that fills only 4–8s ships as a deliberately shorter clip rather than padded dead air.
- Cut-density norms scoped by register (shotlist-director): "1–3 cuts per 15s" is the live-action narrative norm; stylized recipes (3D-animated 6/15s, product montage) are denser by design, cross-linked to style § Style Recipes.
- mm-vs-FOV leak closed (shotlist-director + camera 3.4.1): auto-enrichment lens defaults now speak FOV degrees for Seedance block prompts (63°/47°/29°); the camera lens table carries FOV equivalents per row and flags 45mm-macro as having no FOV anchor.
- FACS micro-beat recipes capped (facs 1.1.1): recipes declared menus, not checklists — pick 2–4 tells per beat; the 3–4-expression cap's logic applies to physical beats.
- HARD RULE 8 word band corrected (root 3.22.1): "500–2,000+" → the actual harvest ladder "218–2,059-word medians".
- Iteration-anchor arithmetic corrected (production-benchmarks): 65–100 generations per kept shot (matches the 1.0–1.5% band; was 50–100), and the "16 finals" figure re-scoped to the most-iterated scene, not per-scene.
- Grey-sheet hex reconciled (ad-asset-prep):
#7f7f7fcreature /#8a8a8ahuman both proven — the rule is pin ONE exact hex per project; added back-links for populated-plate reuse and the canonical Soul ghost-mannequin recipe. - Motion caveat provenance split (motion 3.2.1): catalog stats (~100 presets → ~1,900 variants) attributed to the live catalog pull, not the project harvest.
- Stale 200-word copies given the regime carve-out: troubleshoot 3.0.1 (fix bullet + pre-gen checklist item), audio 3.3.2,
scripts/generate_user_guide.pytip (the v3.22.0 PDF had shipped with the old unqualified rule), and the eval-case description inevals/cases/prompt.json.
v3.22.0
The harvest wave. Absorbs the 13-project community-corpus harvest (2026-07-18: 13 shared Higgsfield projects, 9 creators, ~4,000 production prompts pulled with full params) plus the previously-unported assets from Higgsfield's own skill family (shotlist-builder, seedance-2-pro-director, cinematic-prompt-builder). Chinese-source material re-authored in English.
Added
- higgsfield-seedance 1.11.0:
- Field calibration — the 13-project production corpus
[FIELD]: word-length ladder by register (218w → 2,059w medians — the 50–80w sweet spot is confirmed single-shot-only), register contraction for stylized work, Style Prefix as per-project compiled constant, video briefs hand-authored (enhance_promptoff on video / on for images), observed platform-layer Seedance params (multi_shot_mode: custom,speedramp,bitrate_mode). - Three "helpful-instinct" drift sources with standing locks in POSITIVE LOCKS: environment invention (the #1 drift source, above character drift — "the set contains only what the reference shows"), character-height equalization (heights written into every 2+ character prompt), scale drift on wides.
- Measurable-language additions: masses/sizes in real units for PHYSICS ("50–70 g — it falls gently"), causal prop interaction ("a button press is contact, 2–3 mm travel, click, spring-back — screen lights only AFTER the click").
- Extension prompting: feed the tail, not just the frame — the final 3–4 seconds as
@videoreference carries motion through the join. - Build-safe construction
[OFFICIAL — cinematic-prompt-builder]in the Rewrite Playbook: evacuated-city / energy-standoff / contained-fight / uninhabited-terrain substitutions, containment-doubles-as-physics, the safe benchmark scene. - PRODUCTION-PATTERNS gains a
[FIELD]section: selective motion blur as artifact concealment, HEX-array color lock, generate-forward-reverse-in-edit, populated-plate reuse, off-screen transformation staging.
- Field calibration — the 13-project production corpus
- higgsfield-camera 3.4.0
[OFFICIAL — shotlist-builder]: lens+aperture by shot purpose (85/100mm F1.4 ECU … 45mm macro F2.8) with focus-lock + distortion-forbid clauses, shot-duration-by-type table (0.3–0.5s flash establish → 8–15s full-arc CU), exact-distance micro-move rule (10–15 cm over 7s). - higgsfield-facs 1.1.0
[OFFICIAL — shotlist-builder]: Physical Micro-Beats — the Body Beyond the Face (7 register recipes: throat/breath/skin/posture, incl. suppressed-emotion-as-resistance), anti-AI-video defaults (no tears unless scripted, 0.3–0.5s group-reaction stagger, listeners-in-bokeh are not statues), every line gets three beats (pre/during/post-line), the anti-AI test. - higgsfield-shotlist-director 1.1.0
[OFFICIAL — shotlist-builder + pro-director]: Prompt density — group-when-ALL-5 / split-when-ANY-5 heuristic ("don't fragment grief"), complexity budget + duration ladder, err-toward-more-prompts, auto-enrichment defaults for thin briefs. - higgsfield-style 3.1.0: Register Poles
[FIELD](film vs broadcast-TV vs stop-motion-on-twos vs anime-cel — the style-anchor slot swaps vocabulary by register; one saturated accent reserved for the story) + Style Recipes[OFFICIAL — cinematic-prompt-builder](8 proven shapes: live-action epic, 3D animated 6-shots/15s, game cutscene + pinned HUD, gameplay, FPV oner, product packshot, VFX composite INPUT LOCK, kaiju containment). - image-models.md: Seedream 5.0 Pro (
seedream_v5_pro)[FIELD]— anime/manga sheet + manga-page dialects, art-era anchoring; flagged as absent from the 2026-07-05 spec snapshot (verify live). - production-benchmarks.md: Community-corpus anchors — 13,626 generations for a 2–3-min solo short, TESTS = 61% of the project, 50–100 generations per kept shot (consistent with Hell Grind's 1.0–1.5%), the five-bucket folder discipline, best-second splice culture.
- templates/ad-asset-prep.md: exact-hex grey sheet spec (#8a8a8a)
[FIELD], sibling-face derivation ("spitting image, translated onto…"), ghost-mannequin outfit panel, reverse-angle plates as first-class elements (the 180°-line mechanism), master-plate exposure normalization. - templates/seedance/global-style-prefix.md: Field specimens — the one-axis-per-clause anatomy across the corpus, per-world camera-grammar maps, the reserved-accent discipline, register-aware prefixes; audio policy documented as a project choice, not a law.
- higgsfield-motion 3.2.0: scope caveat
[FIELD]— none of the 13 harvested film productions used a motion preset; presets are the viral-effects product (~100 unique names → ~1,900 per-model variants), film work free-prompts its camera.
Changed
- HARD RULE 8 regime carve-out (root SKILL.md): the 200-word cap now explicitly governs the short-form MCSLA regime only; block-scaffold production prompts replace the cap with structural lint (harvest medians 500–2,000+ words by register).
- Camera-block-at-bottom claim RESOLVED (was "test day pending" since v3.21.0): rejected on field evidence — across ~4,000 harvested production prompts the CAMERA block sits mid-document, never at the bottom. CAMERA-3rd stands.
v3.21.0
v3.21.0 — 2026-07-14
Added
- higgsfield-seedance 1.10.0 — two new Prompt-Craft Laws (2026-07-14):
- Ambiguous verbs — the homograph trap (Peter's field find, covered by no
known prompt guide): if a verb/noun has a plausible second reading
("tearing" = rip vs cry), the model may take it — replace with the phrasing
only one thing can look like; ships with a seed homograph list. - Community v3 cherry-picks (Joey drop, audited vs this skill): camera on
the shadow side + stated operator axis, detail-on-wide "snake cam",
intimate wide, prompt-reset heuristic, canonical-over-plate,
contrast-curve-stated-three-ways. Their camera-block-at-bottom claim is
flagged, not adopted (contradicts CAMERA-3rd; test day pending).
- Ambiguous verbs — the homograph trap (Peter's field find, covered by no
v3.20.1
Changed
- Sora 2 UI presence confirmed (user screenshot of the live model picker, 2026-07-06): the v3.20.0 "verify in the live UI" caveat is upgraded to fact. It is a 4-variant family — Sora 2 (720p) / Sora 2 Pro (1080p) / Sora 2 Max / Sora 2 Pro Max (both 1080p, "BY HIGGSFIELD" enhanced tiers), all 4–12s, multi-shot with sound generation — present in the UI but still absent from the API/MCP catalog (UI generations only). model-guide.md row now carries the variant lineup and real duration/resolution; root SKILL.md and
higgsfield-assist(3.1.1) caveats updated to the confirmed wording.
v3.20.0
Audio specs pipeline + catalog-reality refresh. The specs layer now covers all three output types end-to-end, the dispatcher gained a Load Map, and the two stalest model surfaces (higgsfield-assist, the Sora 2 / "Seedance Pro" mentions) were reconciled with the live catalog.
Added
- Audio specs pipeline:
scripts/sync_specs.py --type audiogeneratesspecs/audio-model-specs.{yaml,json}+specs/AUDIO-MODEL-SPECS.mdfrom the dated audio snapshot (5 models incl.seed_audio);scripts/refresh_specs.pydefault is nowall(video+image+audio —bothkept as the pre-audio alias), audio captured intospecs/cli_baseline.json, tripwire green across all three types. Typed-spec markdown footers now point at their own machine twins (was: everything pointed at the video files).higgsfield-audio3.3.1 cites the generated audio specs. - Load Map (root SKILL.md § Load Map — how much to read): a situation → cumulative-load table so multi-skill loads are deterministic instead of vibes — Fast Path loads two files, everything else adds only what its routing row names.
Changed
- Sora 2 is UI-only — absent from the API catalog and a live
models_exploresearch (verified 2026-07-05). Kept everywhere but annotated: root SKILL.md "What Is Higgsfield?", model-guide.md (video-table row, decision flowchart demoted to parenthetical, camera-control + motion-preset † footnotes, credit table), higgsfield-assist. Root Fast Path default swapped Sora 2 → Seedance 2.0 (action/scale/references). - "Seedance Pro" is a legacy UI label — not in the catalog; annotated in model-guide + assist as superseded by Seedance 1.5 Pro (
seedance1_5) and Seedance 2.0 Fast/Mini (annotate-don't-delete, per the GPT Image precedent). higgsfield-assist3.0.0 → 3.1.0 (first refresh since 2026-04-06): credit-cost tier roster rebuilt against the 2026-07-05 catalog; "Kling 3.0 for anything needing audio" corrected — audio is native across Seedance 2.0/Mini/1.5 Pro (generate_audio), Kling 3.0/2.6 (sound), Veo 3.1 Lite: pick by scene fit, then toggle; plans table now carries a dated verify-live caveat; audio-toggle credit-saving tip added. Kling 2.6 audio cell in model-guide fixed to match the live spec (soundparam, default on).- CS3.5 shot-counter TODO closed (
higgsfield-cinema): API checked 2026-07-05 —multi_promptexposes no maximum and no constraint rule, so the observed cap of 4 is UI behavior the API doesn't document; treat 4 as the working limit, verify live before promising more. - CLAUDE.md: specs/ line + sync command now cover all three types.
v3.19.1
Housekeeping wave: activate the weekly spec-drift schedule, absorb the CLI 1.1.5 release, de-clutter the repo root, and archive the changelog backlog. No prompting-content changes.
Changed
- Spec-drift schedule is LIVE:
HIGGSFIELD_CREDENTIALSrepo secret added; a manualworkflow_dispatchrun verified the full path end-to-end (install → version guard → tripwire → issue-on-drift). The run surfaced CLI 1.1.5, which restores the per-paramenumlists that 1.0.1 dropped (and keepsjob_type+ CEL rules) — the tripwire regains full enum visibility with no parser change. Local CLI upgraded, baseline re-captured with 1.1.5,LAST_VERIFIED_CLIbumped 1.0.1 → 1.1.5, drift issue #78 closed as a cross-version artifact. - Root scripts consolidated into
scripts/: all 9 Python tools (validate.py,build_index.py,sync_specs.py,refresh_specs.py,higgsfield_memory.py,seedance_lint.py,generate_user_guide.py,validate_user_guide.py,sub_skill_descriptions.py) moved viagit mv; internal__file__-derived roots, test/eval importers, CI workflows, slash commands, CLAUDE.md, README, DISCIPLINE.md, and 14 skill files updated. Commands are nowpython3 scripts/validate.pyetc. Root .md/.py clutter roughly halved. - CHANGELOG archived: entries v3.0.0–v3.14.1 (69 releases, ~375 KB) rolled verbatim into
docs/archive/CHANGELOG-v3.0-v3.14.md; root CHANGELOG keeps the current era (v3.15.0+) plus a pointer.
v3.19.0
Seedance 4K Masterclass + Seed Audio 1.0 — content wave from six sources gathered 2026-07-05 (plan: workspace/output/V3.19-PLAN.md): Higgsfield's own downloadable prompt-writter.skill, the official Seedance-4K film tutorial (video + blog, 25 verbatim prompts archived in workspace/input/), spec-verified Seed Audio 1.0 research, a cross-surface video-extension workflow, selected field imports from the community seedance-2.0 repo (v6.6.0), and a character-audition system prompt. Every model claim checked against the fresh 2026-07-05 specs snapshots (shipped in v3.18.1). Provenance tiers used throughout: [OFFICIAL] / [DEMO] / [EMPIRICAL] / [FIELD].
higgsfield-seedance 1.8.2 → 1.9.0
- § Official Prompt Architecture — the Block Scaffold [OFFICIAL — Higgsfield prompt-writter.skill]: the 17-block scaffold (SCENE CONTEXT → POSITIVE LOCKS), distributed-style doctrine (no style prefix), FOV-in-degrees anchor table + CAMERA-3rd-position rule, measurable-language rules (positive-only, km/h, %/meters, human-height scale, left/right-from-camera, Kelvin WB, no director/equipment names), POSITIVE LOCKS, the cut-format ladder + 6-cut vocabulary, tag naming + minimal-reference-text, context isolation, and 4 special protocols (extreme-FOV 4-mechanism stack, whip-pan ≥0.8s, anti-impact locks, observation pattern). Reconciled as a "two regimes" doctrine with the existing six-slot short form; official no-director-names rule noted as overriding the empirical director-substitute trick in block prompts.
- NEW reference
PRODUCTION-PATTERNS.md[DEMO — Seedance-4K film tutorial]: reference-role vocabulary ("100% matches the reference" / "STYLE REFERENCE ONLY, model extends the world" / "VARIETY reference" + the clone-army fix), coordinate blocking (x%/y%, % of frame width, locked screen direction), non-empty opening frame, per-segment LENS LOCK + timed SMASH/MATCH cuts, red-arrow prop annotation, video-reference 1:1 lock + SCREEN REALISM block + duration-match rule, prompted-imperfection realism, 60:30:10 grade, offscreen voice-only characters, specify-what-plays-on-screens, in-prompt scene transitions. - § Extension Prompting — Video-Reference Continuation [EMPIRICAL]: "The scene continues." / "Show me what happens before" openers, occluded-identity
@ImageNbinding, match-source-resolution-AND-duration, chain-degradation + B-roll chain-break, camera-angle-change endings; [FIELD] source-carries-state + references-outrank-text + chain cap ~2 (hard 3) with re-anchor-from-ORIGINAL-references.
higgsfield-audio 3.2.2 → 3.3.0
- § Scene-Audio Generation — Seed Audio 1.0: what it is (one-pass whole-scene audio, released 2026-06-23), decision table vs
text2speech_v2vs Seedancegenerate_audio, verified surface [OFFICIAL — 2026-07-05 audio snapshot] (params, ≤3 audio refs ≤30s XOR 1 image ref,@Audio1..3tokens), script-format prompting clearly labeled [EMPIRICAL — community, NOT official] with a worked example. - Standalone Audio catalog reconciled with the live 5-model catalog (incl. NEW
cozy_voiceengine; game-pipeline-only tools flagged), date-stamped 2026-07-05. - [FIELD] per-language dialogue-sync budget table (EN ~16–20 reliable-sync words per ~15s, Mandarin strongest, RU weak) + voice-reference lip-sync path (rights-sensitive) under Lip-Sync Rules; Supercomputer voice-over pointer [DEMO].
higgsfield-pipeline 3.3.0 → 3.4.0
- § Continuation & Extension Handoff: extend-a-clip workflow, chain management (depth caps + scheduled re-anchoring), source-carries-state rule (stills can't carry motion/camera/audio phase), clean-join planning (angle-change endings, last-channel-on-TV transition trick, post-edit seam note).
higgsfield-soul 3.6.1 → 3.7.0 + asset-prep surfaces
- Two-image character floor (face + full body) + grey-sheet rule [DEMO]; § Split-Panel Outfit-Change Sheet (ghost-mannequin + identity panel); § Variety Sheets — Crowds Without Clones.
templates/ad-asset-prep.md: grey-background canonical home, § Location plates (3/4 angle, empty by default), § Which model makes the sheet (GPT Image 2 4K → Nano Banana Pro on flatness → Soul Cinema for locations/characters); Elements registration.skills/higgsfield-gpt-image-2/reference-sheet-workflow.md: § Views the video will need (front/side/back + BOTTOM/undercarriage for flip shots), § Red-arrow annotation.
higgsfield-character-design 1.0.0 → 1.1.0
- § Screen Test / Audition [EMPIRICAL]: casting read → role options → playable audition lines → voice triggers (≤3 qualities) → final audition prompt (<3500 chars) with divergence rule and a worked mini-example; cross-linked to FACS (direct the takes), Soul (lock the winner), and ad-asset-prep's generate-many → test-in-motion → lock-the-winner loop.
higgsfield-models 3.1.1 → 3.2.0 + guides
- New model rows [spec-verbatim]: Gemini Omni Flash (video), Soul Cast, Soul Location, Nano Banana 2 Lite, OpenAI Hazel (image) across model-guide.md, image-models.md, and higgsfield-models (dual-maintained tables kept in sync); recraft id rename + Seedream 4.5 quality tiers reconciled; GPT Image (original) noted as gone from the 2026-07-05 catalog; utility jobs scoped out with a one-liner.
Root + examples + evals
prompt-examples.md: § Seedance-4K Film Tutorial — Worked Examples — five annotated verbatim excerpts [DEMO] (prompted imperfection, coordinate blocking + LENS LOCKs, 1:1 video reference + SCREEN REALISM, red-arrow lock, VARIETY-reference before/after).- Root SKILL.md: four new routing rows (standalone audio / Seed Audio → audio; extend-continue a clip → seedance + pipeline; asset prep / reference sheets → ad-asset-prep + gpt-image-2 + soul; character audition → character-design). Root version → 3.19.0.
- Evals: +3 cases (Seed Audio scene script, standalone-vs-in-video audio choice, extension continuation) — 43 total.
v3.18.1
Repair wave: un-blind the spec-drift tripwire after the Higgsfield CLI 1.0.1 output-shape change, refresh the specs snapshots (Tier 2, 13 days early — the shape change forced it), and clear the stale-docs debt found by a full repo audit. No new prompting content (that ships in v3.19.0).
Fixed
refresh_specs.pycrashed on CLI 1.0.1 (KeyError: 'job_set_type'— upstream renamed the id key tojob_typeand dropped per-paramenumlists). Now accepts both key generations via_model_id(), and any future output-shape change raises a typedShapeError→ new exit code 4 ("fix the parser") kept distinct from exit 1 ("re-auth") — previously a crash masqueraded as auth expiry inspec-drift.yml. The workflow gained a dedicated exit-4 step and aLAST_VERIFIED_CLIversion guard (warns, never fails, on unverified upstream releases). Fixture-based regression tests added from recorded CLI 1.0.1 JSON (tests/fixtures/cli_1_0_1_*.json); 25 tests intest_refresh.py.- New CEL-rules comparison channel: CLI 1.0.1 ships machine-readable constraint rules (e.g. Seedance's 9-image/3-video/3-audio/12-total reference caps,
mode='fast'forbids 1080p/4k) — exactly the cross-constraint prose the tripwire was historically blind to.cli_view()now captures them; added rules = DRIFT, removed = notice; the channel only compares when both sides carry it (pre-1.0.1 baselines stay silent). sync_specs.pylost the duration envelope: 2026-07 snapshots movedduration_rangeinto adurationparameter (min/max), which silently droppedspec["duration"]and brokeseedance_lint'sduration-out-of-rangerule — caught by thetrap-seedance-overlong-durationeval (the v3.11.2 stale-eval trap working as designed). The envelope is now derived from the duration parameter; parammin/maxare preserved in generated specs.model-guide.mdSeedance 2.0 Mini row: "UI label; not a distinct API id" is no longer true —seedance_2_0_miniis a distinct catalog id (4–15s, 480p/720p, full image/video/audio reference surface, native audio).- Broken link
templates/ad-asset-prep.md→../docs/production-benchmarks.md(file lives at repo root).validate.pynow sweepstemplates/**/*.mdpath-style refs (new[ TEMPLATE PATHS ]section) so this class can't recur. - Stale docs:
photodump-presets.mdimage-snapshot TODO (the snapshot has existed since 2026-06-22); CLAUDE.md counts (30 sub-skills, full templates/ inventory), specs/ description, missing release-gate commands (--strict+ pytest + evals), release-ceremony pointer (tag the merge commit; PDF is an untracked artifact), and stale/project:*slash names (now/validate,/release).release.mdstep 1 now runs all three gates, not just strict validate. DroppedEdit(mnt/**)from.claude/settings.json— it invited exactly the mistake the mnt rule forbids.
Changed
- Specs snapshots refreshed (video + image, 2026-06-22 → 2026-07-05) and a first audio snapshot captured (
models_explore_snapshot_audio_2026-07-05.json:seed_audio= Seed Audio 1.0,sonilo_music,mirelo_text_to_audio,inworld_text_to_speech,text2speech_v2with a newcozy_voiceengine) — machine-readable ground truth for the v3.19.0 Seed Audio content; the audio sync pipeline is a follow-up. - Snapshot deltas absorbed: video +
seedance_2_0_mini/gemini_omni/explainer_video+ utility jobs; image +nano_banana_2_lite/soul_cast/openai_hazel/recraft_v4_1(renamed fromrecraft-v4-1) + utility jobs;seedance_2_0media roles renamed toimage_references/video_references/audio_references(+ keptstart_image/end_image); id renames recorded as spec aliases viaHISTORICAL_IDS(seedance_1_5→seedance1_5,recraft-v4-1→recraft_v4_1, droppedvideo_standarddup) so ledger rows and user inputs written under old ids keep resolving. - CLI baseline re-accepted at 2026-07-05 (tripwire green end-to-end: 0 fresh / 3 change / 1 pull-failed / 4 shape-changed all exercised).
- Root SKILL.md: Shared Resources table now lists all 8
templates/seedance/files and surfacestemplates/ad-asset-prep.md+templates/character-design/; added a consistency tie-break row (higgsfield-soul= lock identity in-platform vshiggsfield-character-design= develop the character first).
Deferred to v3.19.0
- Guide rows/content for the new catalog models (gemini_omni, nano_banana_2_lite, openai_hazel, soul_cast, autosprite, ms_image), Seed Audio 1.0 prompting guidance, and the Seedance 4K masterclass material — content wave, planned at
workspace/output/V3.19-PLAN.md.
v3.18.0 — higgsfield-seedance-vfx (video-to-video footage transformation)
v3.18.0 — 2026-06-30
Video-to-video footage transformation for Seedance 2.0 — take a clip the user already shot, preserve the real subject + camera move, and change one thing (add a VFX element, swap the world, drop in a photoreal creature, relight to match, sync a timed zoom to a line). New 30th sub-skill higgsfield-seedance-vfx. Sourced from a hand-authored practitioner skill + the Higgsfield "Seedance 2.0 in 4K" VFX tutorial (higgsfield.ai/blog/vfx_4k, youtube.com/watch?v=Yte-UGhYkPQ), which demonstrates this exact skill; model-checked against the live seedance_2_0 spec — media_roles include video (v2v input is real), 4k is a legal resolution, plus audio/start_image/end_image, duration 4–15s. One goal-driven release.
New sub-skill — higgsfield-seedance-vfx (1.0.0)
A video-to-video layer on top of higgsfield-seedance (reuses its grammar + preflight linter; only the starting point changes — a real source clip whose subject and camera move must survive). Distinct from the parent's in-clip Transformation prompt mode (a morph generated from scratch). Sections:
- Preserve, then change one thing — lock identity/face/wardrobe/performance/framing/lens/camera, change only the named element, repeat the fragile guardrail at the end.
- Run it in 4K — Seedance 2.0
mode=std4K (faces/lip-sync hold at 4K, warp at 1080p); harmonized with the repo's existing caps (fast → 720p; Cinema Studio → 1080p). The4kenum is model-verified; "detail holds at 4K" is flagged as a practitioner claim. - Prompt anatomy —
@sourcedeclaration, optional@creature/texture reference, specs line (NON-IP guardrail, match-source-runtime,SFXvsSFX and source dialogue only), continuous-shot action, behavioral SFX. - Three levels — L1 swap the world · L2 change an element in-frame · L3 full handheld cinematic (difficulty scales with camera motion).
- Two modes (add an element · replace the environment) + the lighting-integration recipe (color matching alone reads as pasted-in: match key direction, bounce, optics/haze, edges/grounding).
- Photoreal creature integration — biological-accuracy vocabulary (wrinkled/cracked/asymmetric/matte, never smooth/glossy/inflated), telephoto-scale illusion, real contact shadow, reference-image-beats-description.
- Timed camera moves synced to dialogue — dual semantic + numeric anchoring, reveal pull-back with 100%-match landing, lip-sync preservation. Prepended-intro budget arithmetic (
total − intro = surviving window). - Two reference files —
references/dialogue-timing.md(measureT, convert timecodes, phrase both anchors) andreferences/first-frame.md(generate the transformed start still, hand back asstart_image).
Wiring + template
- Routed from root
SKILL.md(routing table + Sub-Skills table),sub_skill_descriptions.py, andINDEX.md. Roster 29 → 30. Root version 3.17.0 → 3.18.0. - New
templates/seedance/footage-vfx-transform.md— fill-in skeleton + four worked patterns (environment swap, head-on-fire, creature-with-reveal-pull-back, full handheld cinematic).
Cross-links (patch bumps)
higgsfield-seedance§ Seedance 2.0 Prompt Modes / Transformation → distinguishes the in-clip morph from v2v footage transform; also added to Related Skills. (seedance 1.8.1→1.8.2)higgsfield-audio§ Related skills → the timed-zoom-to-dialogue + source-dialogue-preservation cases. (audio 3.2.1→3.2.2)higgsfield-camera§ Related skills → preserve a real handheld/driving move frame-for-frame + add a camera move you never filmed. (camera 3.3.0→3.3.1)vocab.md§ Lighting Vocabulary / Scene-physics → the integration recipe for compositing a preserved subject or creature into a new plate.
v3.17.0 — FACS facial-expression control for Seedance 2.0
New 29th sub-skill higgsfield-facs — direct a face by muscle (FACS Action Unit codes like AU12 lip-corner puller, AU6 cheek raiser) instead of emotion labels. The platform's highest-resolution facial control: forced/uncanny/mixed expressions, micro-performance, and honest acting in close-up dialogue.
Provenance: model-checked against the live seedance_2_0 spec, which exposes no FACS/expression field — so the technique is flagged [EMPIRICAL] end to end. The AU vocabulary itself is standard, citeable science; Seedance's interpretation of codes in a prompt is the empirical, not-a-guarantee part.
Highlights
- Plan-first workflow — plan 3–4 expressions → generate a FACS sheet for only those → write the codes. Explicitly counters the "generate the full 49-AU sheet then cherry-pick" anti-pattern.
- FACS reference-sheet generation — parameterized image prompt (GPT Image 2 / Nano Banana Pro) + the LLM-mislabels-AUs caveat (the circulating sheet's
AU82vs standardAU38). - Codes-only vs codes+anatomical-description, the 3–4-expressions-max accuracy limit, optional character photo.
- Full AU code table + EMFACS emotion→AU recipes (Duchenne AU6+AU12, sadness AU1+AU4+AU15, fear, anger, surprise, disgust, contempt).
- Dialogue & monologue facial acting — AU-per-beat schedule combined with the
[AUDIO: Xs]lip-sync block; three worked examples.
Wiring + cross-links
- Routed from root
SKILL.md,sub_skill_descriptions.py,INDEX.md; newtemplates/seedance/facs-expression-beats.md. Roster 28 → 29. - Cross-linked from
higgsfield-soul(Micro-Expressions),higgsfield-seedance(Voice Rewrite),higgsfield-audio(Lip-Sync Rules), andvocab.md(Emotion-as-Visible-Behavior channels).
Verification: validate.py --strict clean · 110 tests pass · 40/40 evals.