Skip to content
AI Primer
workflow

Creators test Seedance reference prompts for longer AI video scenes

Creators refined reference prompts for longer AI video scenes using Seedance last-frame matching, second-by-second timing, character sheets, and GPT Image 2 storyboards. The examples aimed to hold style or character continuity across short shots.

10 min read
Creators test Seedance reference prompts for longer AI video scenes
Creators test Seedance reference prompts for longer AI video scenes

TL;DR

ByteDance's official Seedance 2.0 launch post lists text, image, audio, and video inputs, with up to 9 images, 3 video clips, and 3 audio clips per generation. The coming Dreamina Seedance 2.5 page raises the ambition to 30-second videos and up to 50 references. Creator wrappers are already organizing around that direction: Pollo AI lists Seedance 2.0, GPT Image 2, reference-to-video, and video extension, OpenArt Director pitches chat-based multi-scene direction powered by Seedance 2.0 and GPT Image 2, and VlogMe describes an AI director that turns one idea into scenes, b-roll, lip-sync, audio, and export.

Second-by-second prompts

The strongest Seedance examples used time as the control surface. In techhalla's tortilla sprint prompt, the brief locks the same bald man, beard, yellow TECHHALLA jersey, red dirt location, cooking station, and previous video reference before breaking the action into three timed blocks.

The structure is almost production paperwork:

  • Look: handheld documentary footage, operator shake, harsh high-altitude light, dust, sweat, fabric physics, subtle grain.
  • Continuity: previous video as strict visual and character reference.
  • 0-5 seconds: sprint from the cooking station, toss the tortilla, orbit the airborne flip in slow motion, catch without breaking stride.
  • 5-10 seconds: reach the table, transfer tortilla to bread, add sauce, hand over sandwich.
  • 10-15 seconds: hard cut to the elderly woman taking the bite.

Techhalla used the same grammar in a 1970s Paris cinema verite prompt: film stock, camera shake, period costume, plaza geography, and a 0-3, 3-7, 7-11, 11-15 second walk through the scene.

The funniest detail in techhalla's raw-output reply is that he said the Paris result was raw output. No post-pipeline heroics, just a very specific text-to-video shot list.

The 15-second stitch

Seedance 2.0's public spec fits what creators were working around. ByteDance says the model supports 15-second high-quality multi-shot audio-video output in its official launch post, and GlennHasABeard turned that ceiling into a continuity trick.

Clip two in GlennHasABeard's chase test opens on the exact last frame of clip one, so the edit reads as one chase rather than two stitched generations. Both clips used colored pencil reference images to hold the style across the sequence.

GlennHasABeard later told another creator that the method "definitely unlocks a new level of control" in his reply about the stitch. A separate API note from rainisto said Seedance has 15-second maximums for video and audio, and suggested using the previous 15 seconds as embedded audio plus a 15-second audio reference for the next shot.

Storyboards as shot contracts

AIwithSynthia's VlogMe workflow started with a storyboard, not a finished prompt. AIwithSynthia's storyboard post says the board was made first with GPT Image 2, then brought into VlogMe Video Studio to become a complete cinematic sequence.

The workflow in AIwithSynthia's VlogMe how-to is five steps:

  • Upload the storyboard.
  • Select the video model.
  • Add a prompt.
  • Adjust generation settings.
  • Generate the final video.

AllaAisling split the same idea into a storyboard prompt and a Seedance prompt. The Suburban Male Assembles Furniture storyboard defines eight panels at 1.9 seconds each, then the matching Seedance prompt turns those panels into a 15-second nature documentary comedy.

The same two-layer format appears in AllaAisling's noir pizza short: The Last Slice storyboard frames the empty pizza box, detective, suspects, sauce evidence, cat reveal, and getaway, while the Seedance prompt converts it into 0-1.9 second shot beats.

Character sheets and MCP handoffs

Character sheets moved from prompt garnish to workflow input. magnific's Fable 5 versus Opus 4.8 test used one character sheet, three instructions, and a side-by-side comparison to generate Seedance-ready cinematic scripts.

The setup became explicit in magnific's character-sheet note: give Claude the sheet, then ask it to generate the videos through the Magnific MCP. the Fable 5 script used five timestamped shots, dialogue, music cues, and a button line; the Opus 5 script used a looser five-shot action outline with rain, taiko drums, and specified dialogue.

Claude also showed up as a production coordinator in CharaspowerAI's Odysseus thread. The setup post says one character sheet became a 40-plus-second sequence, with Claude and the Higgsfield MCP turning the myth into shots.

The connector workflow in CharaspowerAI's setup post was: Claude, Settings, Connectors, add Higgsfield MCP, sign in. CharaspowerAI's follow-up said the connector took around 30 seconds to add and under 5 minutes to understand the character sheet, story, and workflow.

Higgsfield's public AI Skills repository lists image, video, 3D, and audio generation across 30-plus models, including Seedance 2.0, GPT Image 2, Veo 3.1, Kling 3.0, Seed Audio 1.0, and Flux 2.

Commercial prompt packs

Food, drink, beauty, and lifestyle prompts got the most detailed product direction. AIwithSynthia's cookies-and-milk prompt even includes a six-panel storyboard infographic before the 15-second ASMR video prompt.

The commercial examples clustered around repeatable prompt blocks:

  • Coffee: AIwithSynthia's iced latte prompt maintains the same woman across a kitchen, street, bakery, and cafe-style ending, with macro shots for ice, milk, espresso, stirring, croissant crust, and ambient sound.
  • Cookies: the cookies-and-milk prompt uses POV hands, doodle callouts, wrapper crinkle, milk pour, cookie dip, cookie break, and a final hero shot.
  • Lipstick: the Higgsfield lipstick prompt uses one reference image to preserve face, hairstyle, makeup, outfit, jewelry, and body proportions during an 8-second beauty spot.
  • Sprite-style drink: AIwithSynthia's beverage prompt locks a reference character, green-and-white palette, can flip, fizz macro, fountains, skate park, and the spoken line "Stay cool. Stay fresh." AIwithSynthia's follow-up repeated the same result in a reply thread.
  • Laundry lifestyle: AIwithSynthia's Sunday laundry prompt stretches an ordinary routine into a neighborhood mini-commercial with sorting, laundromat shots, dryer steam, folding, drawer organization, and ambient audio.
  • Rom-com short: The Wrong Date prompt ran to 90 seconds inside OpenArt Director, with Seedance and GPT handling a mistaken-date setup, cafe wait, accidental meeting, bookstore walk, phone reveal, and sunset handhold.

OpenArt's official Director page uses the same pitch language as these tests: chat through the story and edits, keep faces, environments, voices, and products consistent, and use a structured editor when the scene needs precision.

Genre recipes

Creators were stress-testing tone as much as motion. The best prompts describe the genre camera before the plot.

A few repeatable recipes surfaced:

Tool wrappers

The model was only one layer. The visible workflow moved across model hubs, directors, connectors, and post tools.

The wrappers named in the evidence:

Techhalla's inVideo Agent One thread described a different control loop: generate a batch, review, kill two shots, give notes, run again. his final inVideo step says Agent One animated the locked stills with Seedance 2.0 at 1080p using the structure of his own prompts as reference.

Speed, credits, and longer clips

The most useful caveat came from BLVCKLIGHTai, who had tested the agent workflow three times. BLVCKLIGHTai's critique said inVideo was good but slow, with scripts that would take 4 to 6 hours by hand taking two days through the agent.

BLVCKLIGHTai also said narrative consistency came down to prompting and Seedance, and that gifted credits ran out multiple times while trying to make a 2-3 minute film. In a follow-up on agents, BLVCKLIGHTai ranked inVideo in the top tier but said agents are credit-heavy and conflict with quick iteration while video generations remain slow.

The next platform jump is already being marketed. hasantoxr's Dreamina Seedance 2.5 post said Dreamina Seedance 2.5 is headed for up to 50 multimodal reference assets per generation and up to 30-second continuous video, while Dreamina's official Seedance 2.5 page says the model is coming soon with 4K, 30-second videos, R2V references, and up to 50 references.

Further reading

Discussion across the web

Where this story is being discussed, in original context.

On X· 9 threads
TL;DR2 posts
Second-by-second prompts1 post
The 15-second stitch2 posts
Storyboards as shot contracts5 posts
Character sheets and MCP handoffs5 posts
Commercial prompt packs6 posts
Genre recipes9 posts
Tool wrappers9 posts
Speed, credits, and longer clips2 posts
Share on X