How to make a premium coaching reel from one portrait photo
Give Vizard Agent one portrait photograph and describe what you coach and for whom. Vizard Agent researches you publicly, sources cinematic footage for every scene, generates a portrait scene from your photo, builds luxury typography cards, and filters the caption timings so that no subtitle ever lands across an end card.
What is the short version?
A coaching reel has to look expensive while starting from almost nothing: one photograph and a description of a service. Everything else is sourced or generated, which puts the entire burden on the typography and the pacing rather than on the footage.
- Go to Vizard Agent and upload the portrait.
- Say what you coach, for whom, and in what register.
- Say the length and the call to action.
What do you need before you start?
One good portrait and a clear offer. Vizard Agent researches your name, sources the footage and generates what it cannot find, so this format works from almost nothing. What has to come from you is the offer, because a coaching reel that does not say what it sells is just a mood piece.
- The portrait. One clear photograph is enough.
- What you coach. Fitness, nutrition, mindset, habits.
- Who it is for. It sets the register and the footage.
- The register. Premium, calm, editorial.
- The call to action. How people actually start.
What do you type into Vizard Agent?
Say the register and the audience in the same sentence. Vizard Agent uses that pairing to choose the music, the stock footage and the typography together, and premium coaching for executives looks very different from the same offer aimed at beginners.
Prompt
Variants worth knowing:
- Luxury typography cards. They carry this format more than the footage does.
- A generated portrait scene. So the coach appears in more than one still.
- A clean end card. With the offer and nothing else on screen.
What does Vizard Agent actually do?
Here is the order Vizard Agent worked in on a real coaching reel. Notice how many passes the typography took, and that the last two fixes were both about where text sits on the frame rather than what it says.
- Checks the workspace and locates the uploaded portrait.
- Searches the web for the coach and the service to ground what the reel claims.
- Searches the stock library for cinematic vertical clips and the music library four separate times before finding the right register.
- Analyses the candidate tracks' mood and dynamics, then casts the narrator.
- Generates a portrait scene of the coach from the uploaded photograph and inspects it.
- Records the voiceover lines and checks their durations against the target length.
- Builds luxury typography cards, regenerates them with better fonts, then refines their sizes and visual hierarchy across three review passes.
- Composes and mixes the full soundtrack, recalibrating the loudness twice.
- Renders and reviews, then filters the caption timestamps so no subtitle lands on an end card, repositions the opening title, and moves the header text to the top after checking the frame again.
Step 9 is the detail that makes a premium reel look premium. A subtitle running across a carefully built end card undoes the card entirely, and the fix is filtering the caption track rather than moving the card.
What does the result look like?
From the run this page is written from, probed on the delivered file: 1080x1920, H.264, 30fps, 60.0 seconds, AAC audio. Vertical, exactly sixty seconds, one portrait expanded into a generated scene, cinematic sourced footage, luxury typography cards, narration and captions that stay clear of the cards.
Landing on exactly sixty seconds is the format's own constraint. Vizard Agent composes the soundtrack to that length and fits the narration inside it, because a coaching reel that runs over gets truncated on some placements.
When does this not work well?
This reel makes premium claims about a real person on the strength of a single photograph and a public search. Vizard Agent researches and builds it carefully, and both the claims it makes and the likeness it generates stay your responsibility rather than its own.
- Coaching claims can be regulated. Fitness, nutrition and health advice have rules in many markets.
- A generated scene is not a photograph of you. Do not present it as one.
- Research is thin for private individuals. Vizard Agent grounds what it can find publicly.
- Sourced footage is generic. It sets a register rather than showing your work.
- One portrait limits variety. More photographs give the reel more to cut between.
How do you fix a result that came back wrong?
Name the card or the line. Vizard Agent keeps the research, the generated scene, every typography version, the mixed soundtrack and the filtered caption track, so a repositioned title or a reworded card is a re-render rather than a rebuild.
- "The subtitle covers the end card." The caption timings are filtered around it, as here.
- "The opening title sits wrong." Repositioned and checked on frame.
- "That claim is too strong." Reworded and the narration re-recorded.
How does Vizard Agent compare to doing it yourself?
By hand this means licensing cinematic stock, dropping your one photograph in the middle, and setting typography in a template that was designed for something else. The captions then run over the end card because subtitle tracks do not know cards exist, and that single overlap is what makes a premium reel look like a draft.
| By hand | Vizard Agent | |
|---|---|---|
| Footage for the register | One search, take what fits | Library searched four times for the right mood |
| Typography | A template | Built, regenerated, then refined over three passes |
| Captions over cards | Nobody filters them | Timings filtered so cards stay clean |
| Title placement | Set once | Repositioned twice after checking the frame |
Common questions
Why does this format live or die on typography? Because there is no real footage in it. The stock clips set a mood and the portrait appears once, so the cards carry the actual message, which is why Vizard Agent rebuilt them three times before rendering anything.
How many photos do I need? One good portrait works. Give Vizard Agent more and the reel has more to cut between.
Will it look like me? The generated scene is built from your portrait and is an interpretation rather than a photograph.
How long should a coaching reel be? Sixty seconds, and Vizard Agent builds the soundtrack to exactly that.
Does it research my service? Yes, publicly, to ground what the reel claims. Correct anything it gets wrong.
Why filter the captions instead of moving the cards? Because the cards are the design and the captions are the utility. Dropping the subtitle for the seconds a card is on screen keeps both intact, whereas moving the card to accommodate a subtitle compromises the thing people actually look at.
Can it add my logo? Yes. Vizard Agent places it on the end card alongside the call to action.
Does it narrate? Yes. Vizard Agent casts the voice against the register you name and mixes it into the composed soundtrack.
Does Vizard Agent check the finished reel? Yes. It samples frames across the whole timeline and ran an automated quality check before delivery.