# MiniMax-H3 Multishot - example script (3 shots, 243 frames each)
#
# THE IDENTITY LOCK IS DESCRIPTION DENSITY, NOT ASSERTION.
# Every shot is an independent conditioning pass: the model rebuilds the
# person from the text each time. Writing "same face, same wardrobe"
# asserts continuity without supplying the information needed to rebuild
# it, and the face drifts. Re-describing 6-8 CONCRETE attributes (eye
# colour, hair colour+length+styling, freckles, earrings, specific
# garment and colour) verbatim in every shot is what actually holds it.
# Do the same for the voice: one short concrete line, repeated verbatim.
#
# Timing: 24fps, 243 frames = 10.1s per shot. Speech runs ~2.5 words/sec,
# so each shot wants ~22-25 spoken words. Under-filling a shot is the main
# cause of invented/gibberish speech - the model fills the dead air.
#
# Flowing prose works better than SHOT:/Audio: labels for per-shot scripts.

Clean modern product video, bright and sharp, shallow depth of field: a locked medium shot of a presenter at a white desk in a small studio, warm daylight from a large window on the left, a soft-glow monitor beside her, plants and acoustic panels behind her. She is an attractive American woman in her mid twenties with warm hazel eyes, a friendly confident smile, light freckles, shoulder-length auburn hair tucked behind one ear, small gold stud earrings, and a relaxed sage-green blouse. Her voice is a clear warm young woman's voice in a casual American accent. She looks into the lens and says, "This is multishot for MiniMax H3. One script, one node, and every shot flows into the next — with sound." Her lips move naturally in tight sync with every word. Quiet room tone, her voice clean and close, a faint keyboard click nearby.
---
Clean modern product video, bright and sharp, shallow depth of field: the same locked medium shot of the same presenter at the same white desk in the same small studio, warm daylight from the window on the left, the monitor glowing beside her. She is an attractive American woman in her mid twenties with warm hazel eyes, a friendly confident smile, light freckles, shoulder-length auburn hair tucked behind one ear, small gold stud earrings, and a relaxed sage-green blouse. Her voice is a clear warm young woman's voice in a casual American accent. She gestures easily toward the monitor and says, "Version one point one adds an image start frame, so you can begin from a photo instead of a blank canvas." Her lips move naturally in tight sync with every word. Quiet room tone, her voice clean and close, the soft hum of a computer fan.
---
Clean modern product video, bright and sharp, shallow depth of field: the same locked medium shot of the same presenter at the same white desk in the same small studio, warm daylight from the window on the left, the monitor glowing beside her. She is an attractive American woman in her mid twenties with warm hazel eyes, a friendly confident smile, light freckles, shoulder-length auburn hair tucked behind one ear, small gold stud earrings, and a relaxed sage-green blouse. Her voice is a clear warm young woman's voice in a casual American accent. She leans in slightly, smiles, and says, "It also runs about four times faster on a thirty-two gig card. Grab the workflow, and go make something strange." Her lips move naturally in tight sync with every word. Quiet room tone, her voice clean and close, a light chair shift at the end.
