Vizard Agent

How to keep the same character across every video

Last updated 2026-08-16 · 5 min read

Upload your character to Vizard Agent and describe how they sound and behave. Vizard Agent preserves their exact design across the video, finds a voice that matches the personality you described, and tests that voice before committing to it. On later episodes it works from the same reference so the character stays recognisably themselves.

What is the short version?

A recurring character only works if they are the same character every time. Their face is the easy half — that is a file you keep. The voice is the half that drifts, because "sarcastic" means something slightly different every time somebody picks a preset.

  1. Go to Vizard Agent and upload the character image.
  2. Describe the voice and manner in specific words, not adjectives you like.
  3. Say the format, then keep the same project for every later episode.

What do you need before you start?

One clean image of the character and a description of who they are. Vizard Agent treats the image as an identity to preserve rather than a style to imitate, so what it needs from you in words is everything the picture cannot say: how they sound, how they move, what they would never do.

What do you type into Vizard Agent?

Say "preserve the exact identity" explicitly. Without that clause, Vizard Agent has no way to tell whether your character is a fixed asset or a starting point for something new, and those two readings produce very different videos from the same uploaded image.

Prompt

Create a 9:16 short featuring the uploaded [character] as the only character. Preserve their exact identity, colours and design. Voice: [deep, dry, deadpan, sarcastic]. Keep it under [30] seconds.

Variants worth knowing:

What does Vizard Agent actually do?

Here is the order Vizard Agent worked in on a real character short. More than half the run went on the voice rather than the picture, which is the opposite of what most people expect — and it is the right allocation, because a character whose look is perfect and whose voice is wrong does not read as the same character at all.

  1. Inspects the uploaded character image to understand the design and style.
  2. Checks which generation capabilities are available, then reads the help for the talking-avatar tool specifically.
  3. Reads the voiceover tool's own documentation before choosing settings.
  4. Searches for a voice matching the described personality, then widens the search when the first pass is not close enough.
  5. Generates a test line and listens back to it, judging it against the deadpan, dry register that was asked for.
  6. Reads the voiceover logs when a provider fails, then tries a different named voice rather than accepting a worse match.
  7. Generates the full performance with the pacing and pauses the character needs.

Step 5 is the part that matters. Vizard Agent does not take the first voice that matches the keywords; it makes one and evaluates it, which is what a director does.

What does the result look like?

From the run this page is written from, probed on the delivered file: 720x1280, H.264, 24fps, 28.21 seconds, AAC audio. That run kept 24fps to suit the animation style rather than pushing to 30, which Vizard Agent chose from the character's design.

Later episodes in the same project come back matched to this one, so the specs stay consistent unless you ask for something different.

Expect tens of minutes rather than minutes. This category was not separately measured; comparable work runs a median of 28 to 38 minutes end to end, and across all projects the median cost by tier is Flash 47, Pro 55, Max 242, Ultra 263 credits.

When does this not work well?

Identity holds well when the character is designed to be held: clean shapes, distinctive silhouette, consistent colours. It gets harder as the design gets more detailed, and voice matching is genuinely trial and error rather than a lookup.

How do you fix a result that came back wrong?

Name the drift. Vizard Agent still has the reference image, the chosen voice and the previous episodes, so a correction adjusts the layer you point at rather than regenerating the character from scratch, and the voice you approved stays approved.

How does Vizard Agent compare to doing it yourself?

The manual version is a character sheet, a rigged puppet or a set of poses, a voice actor or a TTS preset you have written down somewhere, and the discipline to use the same ones every time. It works. It also means the consistency lives in your notes rather than in the work.

By hand Vizard Agent
Holding the design A character sheet you refer to The reference image, every time
Finding the voice Audition presets yourself Searches, tests, listens back
Episode twelve Reopen the project, match by eye Same project, one prompt
A detail drifting Spot it in review, fix by hand Name it, Vizard Agent re-checks

Common questions

Can the character speak? Yes. Vizard Agent will find a voice from your description, or use a sample you upload.

Can I use a photo of a real person as the character? You can, and the consent question is yours rather than the tool's. Use yourself, or someone who has agreed in writing.

How many episodes will it stay consistent for? As long as you keep working in the same project. Vizard Agent measures against the earlier work rather than remembering settings.

What if I want to change the character later? Upload the new version and say it replaces the old one. Vizard Agent will match future episodes to the new reference, though earlier episodes obviously keep the old look.

Can the character appear in different scenes? Yes. Describe the scene and Vizard Agent keeps the character fixed while the background changes.