Cindy Zhu.
← all free guides
Content creation Seedance

Make a cinematic AI film of yourself (Tokyo night + the 144p reveal)

Hey, it's Cindy ๐ŸŒฑ The full kit for a cinematic night film of yourself, built in Higgsfield with Soul 2.0 (locks your face) and Seedance 2.5 (films the scenes), that opens like a crunchy 144p game and loads into full quality. Inside: how to lock your face, how to write your own cinematic shots, all nine of my shot prompts, the 144p pixel reveal, and the exact quality-menu green screen from my edit.
the outcome

What you're making ๐ŸŽฌ

Two layers, one video:

The film A nine-shot cinematic sequence of you moving through a snowy city at night: a train window, neon towers, a ramen alley, vending machines, a bridge, a snow run, and one final smile. Same face, same outfit, same colour grade in every shot, generated with Seedance 2.5 inside Higgsfield.
The reveal The opening plays as chunky pixel art, like the video is stuck buffering at 144p. A quality menu clicks up, and the same shot reloads at 240p, then snaps into the real cinematic footage. The pixel versions are made with GPT Image 2 and Seedance.
Everything here is copy-paste. You need one clear reference photo of yourself, Higgsfield for the video generations, ChatGPT for the pixel stills, and CapCut (free) for the edit. Generate everything vertical 9:16.

the build

How it comes together ๐ŸŽฌ

Two engines do the work, both inside Higgsfield. Soul 2.0 locks your face so you look like you in every frame, and Seedance 2.5 turns a written scene into a real cinematic shot. Lock your character once, write each shot, and it films it.


step one

Lock your face first (Soul ID) ๐Ÿง‘โ€๐ŸŽค

The hardest part of any AI film is staying the same person across every shot. Higgsfield gives you two ways to lock it, and knowing the difference is the whole game.

Soul ID: the real lock Train a digital double of yourself once, and every generation after (stills and Seedance videos) is you automatically, no re-uploading per shot. This is the backbone for a multi-shot film. Train on 20 or more varied photos (960px or bigger, different angles and expressions, one full-height shot, no sunglasses). About 3 to 5 minutes and roughly 25 credits.
Reference stills: the quick lock Attach a few photos to one clip. Faster, but weaker across many shots. Use three at most: one front-on, one three-quarter, one profile, all from the same session in the same light. Mixing light or expressions makes the model average them, and averaging is what makes a face drift.
  1. ๐Ÿง‘โ€๐ŸŽค Train your Soul ID (hair up) Build it from your own photos so every shot inherits your face. My locked look is hair up, so I bake that in.
  2. ๐Ÿ–ผ๏ธ Start from a locked frame Generate a perfect Soul still of yourself in the scene first, then animate that exact frame with Seedance image-to-video. Starting from a locked frame beats pure text-to-video for control and burns fewer credits while you dial the shot in.
  3. ๐ŸŽฌ Batch a few takes, cut the best None of this is deterministic. Generate two or three versions of each shot and pick the winner. The edit is what makes separate generations read as one film.
Reference tags to know inside Seedance: @character pins your face, @style borrows a look from a film still, @motion copies the movement of a clip, @audio drives lip-sync or ambience. Scope each one ("use this image for the face only, not its background") so it does not drag in extra people or scenery.

the one rule

Character consistency: the lock block ๐Ÿ”’

Every AI generation is a blank slate. The model has zero memory of the shot you made two minutes ago, so if you describe your character loosely, you get a different woman in every scene. The fix is a lock block: one paragraph stack that pins the face, the wardrobe, the performance, the grade and the lens, pasted at the top of every single shot prompt, word for word. That repetition is what holds five separate generations together as one film.

  1. ๐Ÿ“ธ Pick one reference image A clean, sharp, front-on photo of you in even light. Attach it to every generation. One good reference beats five mixed ones, because the model averages whatever you feed it.
  2. ๐Ÿงฅ Write your own wardrobe lock Swap the outfit paragraph below for what you'll "wear", but keep the same level of detail. Describe clothes as material and construction ("long unbuttoned charcoal-black wool overcoat to mid-calf"), not vibes ("a nice coat"). One strong colour item, like the red scarf here, gives every frame a signature.
  3. ๐Ÿ” Paste it unchanged into every shot The block goes at the top, then the shot description under it. Never reword it between shots.
๐Ÿ”’ the lock block (top of every shot prompt)
CHARACTER LOCK: One woman throughout, matching the attached reference image exactly - same facial
structure, eyes, nose, lips, skin tone, hairstyle, length, colour and parting. Never restyle her hair.
Her face must stay stable, sharp and undistorted in every frame, with no warping or morphing between
shots. Keep the face slim and defined; do not let the model round or soften it, as it tends to drift
toward a softer, younger shape.
WARDROBE LOCK: Identical in every shot. Long unbuttoned charcoal-black wool overcoat to mid-calf. Deep
crimson-red wool scarf wound twice around her neck, one long tail down her front. Black pleated
ankle-length skirt. Cream ribbed socks. Flat black loafers.
PERFORMANCE LOCK: Calm, restrained, melancholic, elegant. She looks like she is remembering something.
Never grin, laugh, pout, point, wave or pose for the camera. Never playful, cute or exaggerated. Emotion
lives in stillness - heavy eyes, slow blinks, a quiet steady gaze. Whenever her face is visible it is
three-quarter or frontal with both eyes visible; never a pure side profile.
GRADE LOCK: Pushed 35mm film stock. Strong halation - every bright practical light bleeds a soft warm
red-orange glow around it, highlights bloom rather than clip. Blacks lifted and milky, faintly
cyan-green, never crushed. Low contrast, gentle S-curve, visible organic film grain heaviest in the
shadows. Cold cyan and petrol-blue ambient, but her face always catches a separate warm 2800K practical
so her skin stays warm and alive, never blue. The crimson scarf is the strongest chroma point in every
frame.
LENS LOCK: 2x anamorphic. Wide oval bokeh, pronounced horizontal blue streak flares across bright
practicals, visible barrel distortion at frame edges, soft corner falloff, shallow depth of field.
Vertical 9:16. Every shot is at night.

the method

Write a great Seedance shot ๐ŸŽฅ

The model reacts to what it can see and measure, not to mood words. "Cinematic", "epic" and "aesthetic" do almost nothing. Translate every one into a physical instruction: a light direction, a lens behaviour, a movement at a stated speed. Build each shot in this order.

  1. ๐Ÿง Subject and one action first Name the subject concretely, then one specific verb in the present tense ("she slowly closes her laptop and looks up"). One action per shot, cramming three in causes morphing.
  2. ๐Ÿ™๏ธ Environment in depth layers Foreground, midground, background, so the shot has real depth instead of a flat plane.
  3. ๐ŸŽฅ Camera third, exactly one move Put the camera line after the subject and environment (at the very end it gets ignored, at the very front it fights the face). Name one move at a speed: "slow dolly-in, 1 to 2 feet", "slow orbit", "locked tripod". Two moves at once is the number one cause of jitter.
  4. ๐Ÿ’ก Light with a direction and a Kelvin The biggest quality lever after the subject. "Warm 3200K practical lamp from the left", "cool neon spill", "backlit rim light". Give light a source and a direction every time.
  5. ๐ŸŽž๏ธ Colour as material plus light Not "red dress, blue wall" (a colour list gets averaged away). Write "crimson silk catching the warm tungsten spill". Add film-look words that land: anamorphic, halation, film grain, shallow depth, soft bokeh.
  6. โฑ๏ธ Beats, length, and the no-cut lock Script beats with time ("at 2s she blinks"), keep one shot to 5 to 8 seconds, and end with "one continuous shot, the camera does not cut on its own" so Seedance stops inventing its own cuts.
Seedance has no reliable negative prompt, and naming a banned thing ("no blur", "no text") can pull it in. Steer positively instead: "sharp, in focus, five fingers, natural skin with visible pores". Add any text or logos as overlays in post, never in the prompt.

Here is a full scene prompt built that way, ready to paste. Swap Tokyo for any world (a rainy Paris street, a desert highway, a spaceship corridor); keep the one continuous take and the slow camera and it stays cinematic.

๐ŸŽฌ a full worked scene prompt
[Reference: my saved Soul character, hair up] A young woman in an oversized cream trench coat walks slowly through a narrow Tokyo backstreet at night, neon signs glowing pink and blue on wet pavement, steam rising from a ramen stall behind her. She keeps walking toward camera, hands in pockets, glancing to the side once. Slow low dolly-back tracking shot, eye-level, shallow depth of field, 35mm cinematic look, gentle handheld sway. Cool night light with warm neon spill, light rain, reflections everywhere. Filmic grain, moody and cinematic, no text. One continuous shot, the camera does not cut on its own. 6 seconds.

part 1

The nine shots ๐ŸŒƒ

Run each one as its own Seedance 2.5 generation in Higgsfield: attach your reference image, paste the lock block, then the shot prompt under it. Each shot is 4 to 5 seconds. The frame next to every prompt is the actual shot from my reel, so you can see exactly what you're aiming for; you can even attach that frame as an extra composition reference alongside your own photo.

The arc matters: she stays quiet and inward for seven shots, and only two shots at the end are allowed to smile. That restraint is what makes the ending land.

Reference frame: she sits by a train window at night, her reflection in the glass, the city streaking past in blue
๐Ÿš† shot 1: the train window (the opening)
MEDIUM CLOSE-UP, three-quarter. Night train interior. She sits by the window on the right of the
vertical frame, body angled toward the glass, the crimson scarf filling the lower frame. Her
reflection floats in the dark glass on the left of frame, softer and dimmer than her real face, the
two faces sharing the frame. Outside the window the city slides past as long horizontal cyan and
teal smears, halating softly. Cool petrol-blue ambient fills the carriage while a warm ceiling
practical keeps her skin warm and alive. She watches the city glide by, eyes soft and unfocused,
lips closed, somewhere else entirely; at 2s she blinks once, slow and heavy, and her gaze drifts a
few degrees to follow a passing light. Locked camera, 40mm anamorphic, faint carriage vibration.
One continuous shot, the camera does not cut on its own. 4 seconds.
Reference frame: low-angle shot of her looking up at glowing billboards, a huge bright screen behind her
๐Ÿ™๏ธ shot 2: looking up at the city (low-angle arc)
LOW-ANGLE MEDIUM CLOSE-UP from below chest height, camera looking up at her against the towers. She
stands in a city plaza at night, a huge bright LED screen glowing white-blue high behind her right
shoulder, dense lit buildings dissolving into large soft oval bokeh all around. Her chin is lifted
and she looks up and around at the glowing signs, eyes tracking slowly from one screen to the next,
lips softly closed, quiet wonder held small. The camera arcs around her at walking pace, 3 km/h, a
single smooth orbital move that keeps her face three-quarter with both eyes visible while the
skyline wheels behind her. Cold cyan ambient, one warm streetlight catching her cheek so her skin
stays warm. 35mm anamorphic, shallow depth, halation blooming off every screen. One continuous
shot, the camera does not cut on its own. 5 seconds.
Reference frame: seen from inside a warm restaurant, she walks past outside in the snow, two diners blurred in the foreground
๐Ÿœ shot 3: past the ramen shop window
MEDIUM WIDE, filmed from INSIDE a warm ramen shop looking out through the window. Foreground: the
dark out-of-focus heads and shoulders of two seated diners frame the lower left and lower right
edges, a wooden counter with bowls between them, the whole interior washed in warm amber. Through
the glass she walks slowly right to left along the snowy alley outside, full figure, passing a wall
of glowing menu lightboxes and red lanterns, snow settled on her shoulders and hair, fresh flakes
falling steadily. She moves at an unhurried stroll, eyes ahead, calm, remembering something; at
2.5s her head turns a few degrees toward the warm window light without breaking stride. Locked
camera, 50mm anamorphic. Warm interior glow against the cold blue street. One continuous shot, the
camera does not cut on its own. 4 seconds.
Reference frame: she sighs in front of a glowing vending machine, her breath visible in the cold air
๐Ÿฅค shot 4: the vending machine sigh
MEDIUM CLOSE-UP, night. She stands before a tall glowing drinks vending machine that fills the
right third of the vertical frame, its pale blue light washing the near side of her face, a single
warm bulb glowing above her. The shot is framed through a dark opening, blurred warm wood
vignetting the frame edges, like the camera is watching from the doorway across the lane. Her face
lifts toward the machine's glow and at 1.5s she exhales a long sigh, the breath rising as slow
visible steam through the cold air, her shoulders dropping a centimetre as it leaves her. Eyes
heavy, somewhere else; at 3s one slow blink. Locked camera with faint handheld drift, 85mm
anamorphic, shallow depth. Cool blue machine light against warm tungsten above. One continuous
shot, the camera does not cut on its own. 4 seconds.
Reference frame: dutch-angle shot of her leaning on a bridge railing, city towers and an orange elevated road behind
๐ŸŒ‰ shot 5: the bridge (she finds the camera)
MEDIUM SHOT, STRONG DUTCH ANGLE, frame canted 30 degrees. She leans on the curved steel railing of
a pedestrian overpass at night, both forearms flat along the rail, weight settled and easy, cream
gloves on her hands. Behind and below her an elevated roadway with an orange-red surface sweeps a
long diagonal through frame, headlights smearing into halating trails; the skyline rises as dense
towers of lit windows dissolving into bokeh. Her face is three-quarter to camera and she looks
directly into the lens, calm and level, holding eye contact for the entire shot; at 2s the wind
moves a strand of hair across her cheek and she lets it stay; at 3.5s the faintest softening
arrives around her eyes, almost a smile that never lands. Slow handheld drift, 35mm anamorphic. One
continuous shot, the camera does not cut on its own. 4 seconds.
Reference frame: she bends to take a red can out of a bright red vending machine in a night alley
๐Ÿฅซ shot 6: grabbing the drink
MEDIUM SHOT from her side, three-quarter, night alley. A bright red vending machine glows on the
left of frame, its white interior light spilling across her. She is bent at the waist toward the
dispenser flap, one hand closing around a cold red can inside it; at 1.8s she straightens back up
to standing in one smooth unhurried motion, the can in her hand, scarf and hair swinging softly
with the movement, and she turns the can once to read it. Behind her the alley falls away into
warm orange lantern bokeh and cool blue shadow, light snow drifting through the streetlight beams.
Locked camera, 50mm anamorphic, shallow depth. One continuous shot, the camera does not cut on its
own. 4 seconds.
Reference frame: motion-blurred handheld shot of her running through snow at night, looking back with a smile
โ„๏ธ shot 7: the snow run (the release)
MEDIUM WIDE, HANDHELD, chasing her at a jog. Night snowfall on an empty tree-lined path, cold
white-blue lamps glowing as soft orbs down the avenue, snow-heavy branches arching overhead, the
whole frame cool cyan. She runs ahead of the camera through fresh snow, coat and pleated skirt
swinging hard with each stride, the crimson scarf tail flying; the camera bounces with live
handheld shake and lags half a beat behind her, the lamps streaking with motion blur. At 2.5s she
looks back over her shoulder at the lens without slowing and breaks into a real unguarded smile,
the one the whole film has been holding back, hair whipping across her face, and keeps running.
Everyone and everything moves at natural real-time speed. 35mm anamorphic. One continuous shot,
the camera does not cut on its own. 5 seconds.
Reference frame: full-body shot of her standing in the snow by a canal railing, looking up at bare branches
๐ŸŒฒ shot 8: the trees (stillness)
WIDE FULL-BODY, locked camera. She stands centered on a snow-covered walkway beside a dark iron
railing, a stone canal wall glowing warm amber behind her, bare snow-dusted branches reaching
across the top of frame. Snow falls steadily through the shot. She stands completely still, weight
even, hands resting at her sides, and looks up into the branches, chin lifted, watching the snow
come down; at 3s her gaze slides slowly back down to the ground in front of her, lashes lowering,
the thought landing. Cream ribbed socks bunched over black shoes, the crimson scarf the only strong
colour in frame. Warm sodium glow against cold blue snow. 40mm anamorphic. One continuous shot,
the camera does not cut on its own. 5 seconds.
Reference frame: close-up of her face with snowflakes in her hair, giving a small soft smile
๐ŸŒฑ shot 9: the smile (the closer)
EXTREME CLOSE-UP on her face, night. Snowflakes rest unmelted in her dark hair and on the crimson
scarf wound up to her chin; large warm golden bokeh orbs float in the darkness behind her. Her
skin detail is crisp and alive, a warm 2800K key on her face against the cool night. She looks
just past the lens, eyes heavy and quiet; at 1.5s her eyes come to the lens and hold; at 2.5s a
small sweet smile arrives, soft and real, barely more than the corners of her mouth lifting and a
light coming into her eyes, and she holds it gently to the end of the shot. Locked camera with
breath-slow drift, 85mm anamorphic, shallow depth. One continuous shot, the camera does not cut on
its own. 4 seconds.
Shots 7 and 9 deliberately break the performance lock: they are the only two moments she is allowed to smile. If you let a grin slip into the first six shots, the ending stops meaning anything. When you run those two prompts, the smile line in the prompt wins over the lock block.

part 2

The 144p reveal ๐Ÿ‘พ

This is the hook of the whole video: it opens looking like it's buffering at 144p, then the quality "loads" up step by step into your real footage. You build it by rebuilding your opening shot as pixel art at two coarseness levels, stills first, then video.

  1. ๐Ÿ–ผ๏ธ Take one clean screengrab Pause your finished opening shot on a readable frame and screenshot it. This one frame is the reference for every pixel step, which is what keeps your subject from drifting position between quality levels.
  2. ๐Ÿ‘พ Make the pixel stills in ChatGPT Upload the screengrab to GPT Image 2 and run the two image prompts below, one per quality step. Replace every [BRACKET] with a short description of your own footage and leave the rest word for word.
  3. ๐ŸŽž๏ธ Make the pixel videos in Seedance In Higgsfield, attach two inputs per step: your original clip (drives all the motion and timing) plus the matching pixel still from ChatGPT (drives the look). Run the two video prompts below.
๐Ÿ‘พ 144p still (GPT Image 2, screengrab attached)
Redraw this image as the SAME pixel-art scene at ONE small quality step lower - like the
same picture shown at 240p instead of 360p on YouTube. It must stay clearly recognizable:
[YOUR SUBJECT] and the key background elements are all still obvious. This is a subtle
reduction, NOT a heavy pixelation.
WHAT CHANGES (keep it gentle): pixels get modestly larger, roughly 1.3 times bigger blocks -
no more than that. Slightly fewer colors and a touch less fine detail. Merge only the very
smallest details into their neighbors. Everything else stays intact and readable.
WHAT STAYS THE SAME: identical composition and framing, same colors and lighting as the
input, [YOUR SUBJECT] in the same position, same background elements ([LIST YOUR BACKGROUND,
e.g. "trees, buildings, sky"]). Same mood as the input.
STYLE: keep it clean flat pixel-art on a uniform grid - solid color blocks, hard clean edges,
no painterly shading, no gradients. Just a slightly coarser version of the input's own style.
SUBJECT: [DESCRIBE YOUR SUBJECT - its color, shape, and any defining feature that must NOT
change. If it includes text, a face, or a license plate, say it should stay blank/unreadable.]
DO NOT: over-pixelate, go 8-bit or retro, make it blocky or abstract, lose recognizable
shapes, blur, smudge, add gradients, change colors or lighting, or alter your subject's
defining features. A subtle, clean, slightly-lower-res version only. Match your original
aspect ratio.
๐Ÿ•น๏ธ 240p still (GPT Image 2, screengrab attached)
Redraw this photo entirely as a CLEAN pixel-art video game render - the crisp, flat
cel-shaded look of a stylized mobile game screenshot. This is a full re-illustration, NOT a
filter, blur, or downscale. Rebuild every element from crisp, hard-edged square pixels on a
strictly UNIFORM grid - every pixel the same square size.
STYLE: flat solid single-color fills with clean cel-shaded blocks - no painterly shading, no
gradients inside shapes, no photographic texture. Bold simple shapes like game assets. Crisp
and clean.
COLOR AND LIGHT (IMPORTANT): match the ORIGINAL photo's colors and lighting faithfully -
[DESCRIBE YOUR LIGHTING, e.g. "golden-hour sunlight", "overcast daylight", "neon night
lighting"]. Translate these exact colors into flat pixel blocks. Do NOT brighten it into
midday, do NOT make it dark or moody either - the same mood as the photo, just rendered as
clean pixel art.
KEEP THE SAME: exact composition and framing, [YOUR SUBJECT] in the same position, [DESCRIBE
YOUR BACKGROUND LAYOUT - what's left, right, and behind your subject].
SUBJECT: [DESCRIBE YOUR SUBJECT'S DEFINING FEATURES AND WHAT MUST NOT CHANGE.]
TEXT / FACES / PLATES (if relevant): keep any readable text, license plates, or faces blank
or unrecognizable - no readable letters, numbers, or identifying detail.
PIXEL SCALE: a fine, crisp, clean uniform grid - detailed but obviously an illustrated game
render, not a photo. Match your original aspect ratio.
DO NOT: overbrighten into midday, use dark GTA-style moody shading, painterly brushwork, soft
gradients, blur, film grain, mixed pixel sizes, or a muddy grid. Do NOT alter your subject's
defining features. No readable plate/text, no watermarks, no UI overlays.
๐Ÿ‘พ 144p video (Seedance 2.5, original clip + 144p still attached)
Completely REDRAW this footage as authentic retro video-game PIXEL ART, in the style of
an early 1990s console game (NES / Game Boy Color / early Super Nintendo, like Mega Man or
Street Fighter II). Video 1 controls ALL motion, camera movement, timing and composition -
keep them EXACTLY; only the rendering style changes. The ENTIRE frame must sit on ONE uniform,
coarse square-pixel grid - background, [YOUR SUBJECT], and the setting all rendered at the
same large pixel size. Big chunky visible square pixels. Hard-edged flat color blocks with a
small limited palette (roughly 16 colors), ordered dithering for any gradient, thick clean
outlines, no smooth gradients, NO blur, NO soft edges, NO painterly brush texture. Crisp
sprite look. Redraw [YOUR SUBJECT] cleanly as a game sprite at the same pixel size as
everything else, keeping its defining features ([DESCRIBE, e.g. "stays dark, never green"]).
Keep the same lighting and layout as the source. Do NOT add characters, text, HUD, or UI.
Ignore the source audio.
๐Ÿ•น๏ธ 240p video (Seedance 2.5, original clip + 240p still attached)
Completely REDRAW this footage as authentic 16-bit Super Nintendo PIXEL ART (SNES /
Neo Geo era, like Street Fighter II or Metal Slug). Video 1 controls ALL motion, camera
movement, timing and composition - keep them EXACTLY; only the rendering style changes. The
ENTIRE frame must sit on ONE uniform square-pixel grid - background, [YOUR SUBJECT], and the
setting all at the same pixel size. Clearly visible square pixels, but slightly finer and with
a richer (but still limited) palette than 8-bit. Hard-edged flat color blocks, light dithering,
clean outlines, crisp sprite detail, NO blur, NO soft edges, NO painterly texture. Redraw
[YOUR SUBJECT] as a clean sprite at the same pixel size as the scene, keeping its defining
features ([DESCRIBE, e.g. "stays dark, never green"]). Keep the same lighting and layout as
the source. More detail and depth than 8-bit but still an obvious retro pixel-art sprite
scene, not photoreal. Do NOT add characters, text, HUD, or UI. Ignore the source audio.

the overlay

The quality-menu green screen ๐Ÿ“ผ

The thing that sells the illusion: a YouTube-style quality dropdown that clicks from 144p up through 240p and 720p to 1080p, on a solid green background. This is the exact overlay from my edit, yours to download:

Green screen overlay preview: a video quality dropdown menu listing 144p, 240p, 720p and 1080p with a cursor about to click

A 6-second 1080x1920 clip: the menu opens, the cursor picks a quality, the menu closes.

๐Ÿ“ฅ Download the menu green screen (MP4, 0.6 MB)

  1. ๐ŸŽฌ Drop it over your edit in CapCut Add it as an overlay layer above your main timeline.
  2. ๐ŸŸฉ Key out the green Tap the overlay, then Cutout, then Chroma key, and pick the green. Only the menu and cursor stay.
  3. โฑ๏ธ Time the click to the jump Line the cursor click up with the exact frame your footage cuts to the next quality step. Scale and position the menu wherever it reads best.

assembly

Cutting it together โœ‚๏ธ

  1. ๐Ÿ“ถ Lay the clips low to high 144p pixel video, then 240p, then your real cinematic footage, ending at full quality. The reveal only works in one direction.
  2. โšก Keep each step short and cut hard Roughly 0.3 to 0.6 seconds per pixel step, hard cuts on the beat. A dissolve kills the loading feel instantly.
  3. ๐ŸŽต Cut on the music Put each quality jump on a beat and let the track open up as the image sharpens. The sound doing the same "unlock" as the picture is what makes people rewatch.

troubleshooting

Fix the most common Seedance problems ๐Ÿ”ง

Face changes between shots Lock it with Soul ID, or attach the same three matched stills every time, and repeat one short identity line word for word. Do not pile on new facial detail, that makes drift worse.
Face morphs mid-shot Simplify the scene and the action so the model spends its capacity on the face. On 2.5 you can region-edit just the face instead of re-rolling the whole clip.
Plastic, over-smooth skin Ask for "natural skin texture, visible pores, film grain", and add a touch of grain to your reference photo first. A too-clean reference reads as plastic.
Jitter or flicker Too much at once. Slow the camera to one move, simplify the action, shorten the shot.
It adds its own cuts End the prompt with "one continuous shot, the camera does not cut on its own", and name the perspective ("single continuous shot, eye-level").
Warped hands or stray text Keep hands in slow, clear gestures ("five fingers visible"), and never ask for on-screen text, add captions in CapCut.
Prompt bloat is the real enemy: past roughly 150 words the model starts silently dropping instructions. Keep the scene text short, the identity line tight and repeated, and cut decorative adjectives. If an artifact survives two takes, stop adding fixes and change strategy: reword the subject or simplify the camera.

finish sharp

Resolution, takes, and credits ๐Ÿ“ถ

  1. ๐Ÿงช Prototype at 720p, silent Validate motion and framing cheaply. Resolution and audio are the two big credit sinks, so add them only on the keeper.
  2. ๐Ÿ” Batch takes, then lock Generate a few versions per shot, pick the best, then commit to the final render. An unlimited plan is the sane way to fund "many takes".
  3. ๐Ÿ” Upscale last Finish the locked cut, then upscale to 1080p or 4K (Higgsfield has Topaz built in). Do not blind-render finals at 4K, it is expensive and you will re-roll.
Seedance 2.5 is the current model (longer takes, better face lock, sound in one pass). It renders around 720p then upscales, so if your account still shows Seedance 2.0, these shots work there too; 2.0 just renders higher natively over a shorter take.

the craft

6 rules that keep it cinematic ๐ŸŽฏ

Keep the lock block identical Rewording it between shots is the fastest way to lose the face. Copy, paste, never retype.
Face drifting? Cut scene text Move the lock block to the very top and shorten the scene description. Adding more facial detail makes drift worse, not better.
Restraint is the whole look Heavy eyes, slow blinks and stillness read cinematic. A grin, a wave or a pose for the camera breaks the mood instantly.
Generate location siblings together When two shots share a place, generate one first and feed the result back as an extra reference so both read as the same real location.
Grade in the stills, not the video Colour-correcting a still frame is cheap. Correcting it in video means regenerating the whole shot.
Take several takes per shot Generate 2 or 3 versions of each shot and cut the best. The edit is what makes separate generations read as one film.

The links ๐Ÿ”—

๐ŸŽฅ Higgsfield (Soul 2.0 + Seedance 2.5 live here): higgsfield.ai
๐Ÿค– Claude (your prompt enhancer, a rough idea into a full scene prompt): claude.ai
๐Ÿ‘พ ChatGPT (GPT Image 2 for the pixel stills): chatgpt.com
โœ‚๏ธ CapCut (free, for the edit + chroma key): capcut.com
๐Ÿ“ผ The quality-menu green screen: download the MP4

Follow @cindiezhu for more AI tips every single day ๐ŸŒฑ