Vizard Agent

How to cut a long video into short clips that end on a finished sentence

Last updated 2026-08-16 · 6 min read

Give Vizard Agent the video and a maximum clip length. Vizard Agent transcribes it, finds candidate passages that fit inside your limit, then verifies every boundary against the word timings and adjusts any that would have cut a sentence in half — so each clip starts and ends on a complete thought.

What is the short version?

Cutting a long video into short clips is trivial if you cut on the clock and useless if you do. A twenty-second clip that stops halfway through a sentence is unpostable, and the fix is to choose the boundaries from the speech rather than from a timer.

  1. Go to Vizard Agent and give it the video or a link.
  2. Say the maximum length each clip may run to.
  3. Say how many you want, or let Vizard Agent choose the strongest passages.

What do you need before you start?

The video and a limit. Vizard Agent finds the passages itself from the transcript, so the numbers you supply — the ceiling on length and roughly how many clips you want — are the only real constraints it needs before it starts looking for places to cut.

What do you type into Vizard Agent?

Give the ceiling and the source. Vizard Agent works out where the natural boundaries are on its own, so there is no need to nominate timestamps — and a single sentence naming the limit is genuinely the whole brief for this job.

Prompt

Cut this video into several clips, each no longer than [20] seconds. [link]

Variants worth knowing:

What does Vizard Agent actually do?

Here is the order Vizard Agent worked in on a real split. Half the run is spent checking boundaries it has already chosen — and that checking is what caught a clip that would have ended in the middle of a sentence.

  1. Downloads the video from the link.
  2. Analyses its metadata and transcribes the audio to find natural cut points.
  3. Locates the transcript file and inspects its structure and keys.
  4. Finds high-quality, non-overlapping candidate passages that fit under the limit.
  5. Confirms the segments carry word-level timings for precise trimming.
  6. Extracts exact word-aligned start and end times for all five clips.
  7. Adjusts the end of the second clip because it would have cut off mid-sentence.
  8. Verifies the boundaries of every other clip in turn, printing the surrounding lines to check each one.
  9. Cuts them at exact frame boundaries, then extracts a frame from each to confirm the picture is right.

Step 7 is the whole point of this page. Vizard Agent had already chosen its boundaries, went back to check them one by one, and found one that needed moving — which is a habit rather than a feature.

What does the result look like?

From the run this page is written from, probed on a delivered clip: 1920x1080, H.264, 30fps, 20.0 seconds, AAC audio. Exactly at the limit, in the source's own shape, starting and ending on complete sentences — one of five clips cut from a much longer recording in a single pass.

Each clip is uploaded with its own link, so they arrive as five separate files ready to schedule rather than as one video with markers in it.

This category was not separately measured, so treat the timing as a range rather than a promise: comparable work runs a median of 28 to 38 minutes end to end. Across all projects the median cost by tier is Flash 47, Pro 55, Max 242, Ultra 263 credits.

When does this not work well?

The transcript is what makes the boundaries good, so anything the transcript does not capture is invisible to the selection. Vizard Agent verifies every cut point against the word timings, and a clip can still be technically clean and contextually meaningless.

How do you fix a result that came back wrong?

Name the clip. Vizard Agent keeps the transcript, the word timings and the chosen boundaries, so extending a clip's opening or replacing a weak one is a re-cut from measurements it already holds rather than a second pass over the video.

How does Vizard Agent compare to doing it yourself?

By hand this is scrubbing for passages, setting in and out points, exporting, and then discovering on playback that two of the five stop mid-word. The cutting itself takes minutes; the checking is the part people skip, and it is the part that decides whether the clips are usable.

By hand Vizard Agent
Finding passages Scrub and guess Candidates from the transcript
Setting boundaries Snap to the timeline Word-aligned start and end
Checking each one Watch all five back Boundaries verified, one adjusted
Delivering them Export five times Cut and uploaded with links

Common questions

Will the clips cut off mid-sentence? No. Vizard Agent sets the boundaries from word timings and re-checks each one before cutting.

Can it work from a link? Yes. Vizard Agent downloads the video and works from the file.

How many clips will I get? As many good passages as fit inside your limit, or exactly the number you ask for if you name one. Vizard Agent selects rather than slicing on the clock.

Can they be reframed for a phone? Yes. Ask for vertical and Vizard Agent re-frames on the subject through each clip rather than cropping the middle of the frame and hoping the speaker stays in it.

Can it add captions to each one? Yes, and cheaply, since the word timings already exist from the transcription.

What if the video has no speech? Then there are no spoken boundaries to find, and asking Vizard Agent to cut on shot changes instead is the better request.

Do the clips arrive separately? Yes. Vizard Agent delivers each one as its own file with its own link, ready to schedule.