How to rank the best moments of a video into a countdown short
Upload the video and ask Vizard Agent for a ranked countdown of its best moments. Vizard Agent maps the shots, transcribes the dialogue, picks and orders the extracts, blurs any existing captions at their measured position, labels each entry with its rank, and checks the loudness of the finished render.
What is the short version?
A ranking gives a compilation a reason to be watched all the way to the end. Numbering the moments turns "here are five funny bits" into "the funniest one is still coming", and that promise is the entire mechanism the format runs on.
- Go to Vizard Agent and upload the source video.
- Ask for a ranked countdown of the best moments.
- Say if there are existing captions that should be hidden.
What do you need before you start?
The video and a criterion to rank by. Vizard Agent finds and orders the moments itself, but "funniest", "most surprising" and "most useful" produce genuinely different rankings out of exactly the same footage, so the word you choose matters.
- The video. One source is enough.
- The criterion. Funniest, best, most surprising.
- How many. Five is the usual number.
- What to hide. Existing subtitles or watermarks.
- The frame. Vertical for a short.
What do you type into Vizard Agent?
Name the criterion and the number of entries. Vizard Agent builds the ranking from that, and mentioning any existing captions in the same request matters because they have to be located before the extracts are cut rather than patched afterwards.
Prompt
Variants worth knowing:
- Counting down or up. Ending on the best either way.
- Ranking cards between entries. Rather than a corner label.
- Existing captions removed. So your own can be added cleanly.
What does Vizard Agent actually do?
Here is the order Vizard Agent worked in on a real ranked short. The step that keeps the result clean is the tenth: it located the exact position of the existing subtitles before deciding anything about the blur.
- Watches the source video to spot the strongest passages and the on-screen text.
- Downloads the source and prepares a visual overview to understand its structure.
- Loads the ranking style before cutting anything.
- Builds a contact sheet to see the whole video quickly.
- Runs the shot analysis and the transcription to locate the passages.
- Checks the verified effects for an opening and for transitions.
- Reviews the contact sheet to identify the scenes and the faces.
- Transcribes the dialogue with timings and maps the shot changes.
- Verifies the chosen extracts and the exact position of the existing subtitles, renders the vertical short with the ranking labels and the blur, fixes the filter graph, places the loudness normalisation correctly and measures the final render's integrated level.
Step 9's ending is the discipline. A ranking short cuts between extracts recorded at different levels, and measuring the finished file rather than trusting the mix is what stops entry four being twice as loud as entry three.
What does the result look like?
The source this was built from measures 1910x1060 at 30fps and runs 266.03 seconds. From it comes a vertical short in which the chosen moments appear in ranked order with numbered labels, the original subtitles blurred out beneath them, one consistent loudness across every extract, and an opening built to set up the countdown.
The ranking is the structure. Without it the same five clips are a compilation somebody stops watching after the second one, which is why the labels matter more than they look.
When does this not work well?
A ranking implies a judgement, and that judgement is only ever as good as the criterion you handed to Vizard Agent in the first place. These are the situations where the format struggles no matter how carefully the cutting is done.
- A vague criterion gives a vague order. "Best" means less than "funniest".
- Similar moments flatten the ranking. Five equivalent clips have no top.
- Blurring leaves a visible patch. Existing captions do not vanish invisibly.
- Levels vary between extracts. They have to be measured, not assumed.
- Rankings are opinions. Say so if it matters to your audience.
How do you fix a result that came back wrong?
Say which entry is in the wrong place and where it belongs. Vizard Agent keeps the contact sheet, the transcript, the shot map, the subtitle positions and the loudness measurements, so the order can change without recutting the extracts.
- "Number two should be number five." Re-ordered and the labels rebuilt.
- "The blur is obvious." Repositioned from the measured subtitle location.
- "Entry three is too quiet." Re-levelled from the measured render.
How does Vizard Agent compare to doing it yourself?
By hand this means watching a five-minute video through several times over to rank its moments properly, masking the existing subtitles on each extract separately, and then balancing five clips recorded at quite different levels entirely by ear.
| By hand | Vizard Agent | |
|---|---|---|
| Finding moments | Watch repeatedly | Shot map and transcript |
| Ranking | Decide from memory | Ordered against a stated criterion |
| Existing captions | Mask each extract | Position located once, applied throughout |
| Levels | Balance by ear | Integrated loudness measured on the render |
Common questions
How many entries work best? Five. Enough for Vizard Agent to build a countdown, short enough that nobody leaves early.
How does Vizard Agent rank them? Against the criterion you name. Different criteria produce genuinely different orders.
Can it hide subtitles that are already burnt in? Yes. Vizard Agent locates their exact position first and blurs or covers that region.
Should it count up or down? Down, usually. Either way Vizard Agent ends on the strongest entry.
Can I choose the moments myself? Yes. Name them and Vizard Agent handles the ranking, labels and levelling.
Why measure the loudness of the finished file? Because a ranking short is built from extracts recorded at different levels, and a mix that looked right while building can still deliver one entry noticeably louder than the rest.
Does it work on any kind of video? Any with distinguishable moments. Gaming, interviews, compilations and reactions all fit.
Can I get a longer version? Yes, with more entries, from the same shot map and transcript.
What if the moments are all similar? Then the ranking is arbitrary and Vizard Agent will say so rather than inventing an order.
Should the ranking be stated at the start? Announcing the count helps. Vizard Agent builds an opening that sets up how many entries are coming, because a viewer who knows there are five will stay for the fifth.
Can the same source give several shorts? Yes. Vizard Agent already holds the shot map and the transcript, so a second ranking on a different criterion costs a fraction of the first one.
How long should each entry run? Five to eight seconds for most formats. Vizard Agent keeps them tight because the countdown itself is doing the work of holding attention, and a long entry lets the viewer off the hook before the next number arrives.
Does Vizard Agent check the finished short? Yes. It reviews the framing, the labels and the blur across the whole render.