How to make the next scene continue exactly from where the last one stopped
Describe the exact state the previous scene ended in and what happens next. Vizard Agent builds a character and scene reference first, generates every keyframe anchored to it, and then produces the clips — so the set, the props and where everyone was standing carry into the new scene along with the faces.
What is the short version?
A series that runs as one continuous story has a harder job than one with separate episodes. The characters have to look the same — that is the well-known part — and so does the room, the weather, the pile of blocks in the road and who was standing where.
- Go to Vizard Agent and describe where the last scene ended.
- Say what must not change: the set, the cast, the style.
- Say what happens next, as one beat forward.
What do you need before you start?
A description of the ending state, and the previous video if you have it. The state matters more than the footage — "the six vehicles are waiting at the entrance of the tunnel" is a more useful starting point than a frame, because it says what the next scene has to honour.
- The ending state. Where everything was left.
- The cast. Named, with their look described.
- The set. The place, and what is in it.
- The style. In the same words every time.
- The next beat. One step forward, not a whole episode.
What do you type into Vizard Agent?
Write the ending state as a fact, not as a memory. The phrase that carries this is "directly from the previous scene, where…" — it tells Vizard Agent that the opening is not a fresh establishing shot but a continuation of a situation already in progress.
Prompt
Variants worth knowing:
- "Directly from the previous scene." The instruction that binds it.
- "Keep the exact same set." Not just the same characters.
- "One beat forward." Stops it compressing a whole arc.
What does Vizard Agent actually do?
Here is the order Vizard Agent worked in on a real continuing animated series where each video picks up from the last. Steps two and three are the whole mechanism: nothing else gets generated until there is one single reference image for every later scene to agree with.
- Checks the generation tools and loads the film approach.
- Generates one character and scene reference image.
- Looks at that reference before building anything on it.
- Generates the scene keyframes anchored to the reference.
- Regenerates them against the reference URL for consistency.
- Tiles the keyframes into a grid and reviews them together.
- Generates the clips for every scene concurrently.
- Probes the clips and builds a frame contact sheet.
- Reviews the contact sheet across the whole continuation.
- Analyses the motion timing of each clip.
- Builds the audio — beds and action-matched effects.
- Tests the framing on each clip and checks the crops.
- Renders the cut with the sound aligned.
- Measures the loudness and renders the master.
- Builds a waveform and timeline check frames and reviews both.
- Renders a clean version with plain scene cuts as an alternative.
- Analyses the finished film for sync and mix.
Step two is what people skip. Generating the new scene straight from a written description gives you the same characters described, which is not the same as the same characters shown — the reference image is what makes "the same" checkable rather than hopeful.
Step six matters for continuity specifically. Reviewing the keyframes as a grid, before any motion exists, is where a wrong-coloured building or a missing prop gets caught — while it is still cheap.
What does the result look like?
A scene that opens where the last one closed, with the same cast in the same place, and moves one beat forward from there. Nothing re-establishes itself; a viewer watching them back to back sees one continuous piece.
The consistency covers the set as well as the faces, which is the part that usually slips. A character who looks right standing in a room that has quietly changed colour reads as a mistake just as loudly.
When does this not work well?
Continuity degrades over a long run of scenes. Each one is generated against a reference, and small drifts accumulate across many of them unless Vizard Agent refreshes that reference periodically rather than inheriting a copy of a copy.
- Very long series. Drift compounds; re-anchor periodically.
- Complex sets. More detail means more to get subtly wrong.
- A beat that is really three. Compress it and continuity breaks.
- Vague ending states. "They were outside" is not enough.
- Live-action footage. This is a generation technique, not a shoot.
How do you fix a result that came back wrong?
Name the specific thing that changed between the two scenes. Vizard Agent keeps the reference image, all the keyframes and the scene descriptions, so a corrected prop or a wrong colour re-anchors only the affected scene rather than the whole series.
- "The tunnel is a different colour." Re-anchored to the reference.
- "They start somewhere else." The opening state re-stated and rebuilt.
- "Too much happens." Cut back to one beat.
- "The style drifted." The reference refreshed and re-applied.
How does Vizard Agent compare to doing it yourself?
By hand, continuity in a generated series is a discipline: keep a reference, describe the state, check every scene against both. It works and it is entirely on you to remember — and the scene that breaks it is always the one made in a hurry.
| By hand | Vizard Agent | |
|---|---|---|
| The reference | If you kept one | Built first, then anchored to |
| Checking before motion | Rarely | Keyframes reviewed as a grid |
| Set continuity | Often forgotten | Held with the characters |
| Catching a drift | On playback | Before the clips are generated |
| The next scene | Start over | Continues from the stated ending |
Common questions
Do I need the previous video? It helps, but Vizard Agent works from a clear description of the ending state.
What is the ending state? Where everyone and everything was left. Vizard Agent opens the next scene there.
Will the characters look the same? Yes. Vizard Agent anchors them to one reference image.
What about the set? Vizard Agent holds it too, which is the part that usually slips.
How many scenes can it continue? Several before drift shows. Vizard Agent can re-anchor when it does.
Can I change one thing deliberately? Yes. Say what changes and Vizard Agent keeps everything else fixed.
Is this the same as making the next episode? No. An episode restarts the situation; Vizard Agent picks up inside it.
Can it do live action? Not this way. Vizard Agent is generating here, not shooting.
Will the audio carry over too? Yes, if you ask. Vizard Agent keeps the sound treatment consistent.
Can I get a version with plain cuts? Yes. Vizard Agent renders an alternative without the transitions.
What if the reference is wrong? Fix it once. Vizard Agent rebuilds everything anchored to it.
Does it check before rendering? Yes. The keyframe grid is reviewed before any clip is generated.
Can I keep a bible of the cast? Yes, and it helps. Hand Vizard Agent the same description each time.
Why build a reference first? Because "the same characters" written down twice produces two different characters, and one image produces one.