Review

Vizard AI Model v1 vs v2: Which Is Better?

We ran the same video through Vizard AI's Model v1 and the new v2 on the free plan. v1 gave 10 clips, v2 gave 5 — and v2 costs 30% more credits.

Disclosure: CreatorIntelHQ may earn a commission if you buy through some links. This page is based on hands-on testing and documented observations.

TL;DR: Vizard added a second clipping model, Model v2, on 8 July 2026. Both are available on the free plan — v2 is not paywalled. On an identical 10-minute video, v1 returned 10 clips and v2 returned 5, with v2's clips running roughly 1.8x longer. v2 costs about 30% more credits. Captions and export quality are identical between them, so the premium buys clip selection — nothing else.

This page answers one question: when you see that "Model v1 / Model v2" dropdown on Vizard's upload screen, which one should you pick? If you want the full tool walkthrough instead, read our Vizard AI review. Here we stay narrow — same video, same settings, both models, and the numbers that came back.

By Gaurav Sharma · Last updated August 25, 2026 · Evidence status: Strong

Based on hands-on testing · How we test creator tools

What We Tested

We ran Vizard exactly as a creator on the free plan would, with no paid features unlocked.

The setup:

  • A fresh free-plan account with 60 monthly credits
  • One source file: a 10:01, 1920x1080, 30fps video, uploaded directly (not pasted as a link)
  • Identical settings on both runs — 9:16 ratio, "Any length" clips, Default template, Add emojis on, Highlight keywords on, B-rolls off, Remove silences off, Auto-censor off
  • No subtitle file attached, so captions had to come from Vizard's own transcription
  • Credits recorded before and after every step
  • Output clips downloaded and inspected with ffprobe

The source was a NASA Artemis II mission broadcast — public domain, two presenters at a desk, continuous speech throughout, no burned-in captions. We use the same file across every tool we test so results stay comparable.

One thing we could not control: the v2 run had spoken language set explicitly to English, while the v1 run was left on the default. For unambiguous English audio we do not think this changed anything, but it means this is not a perfectly controlled A/B and we would rather say so.

What Each Model Costs

Vizard prices by the minute of source video, and v2 carries a premium:

Source length Model v1 Model v2
10:01 10 credits 13 credits
22:48 22 credits 28 credits
Effective rate ~1.0 credit/min ~1.3 credits/min

Vizard AI import screen showing the Model v1 and Model v2 dropdown with credit costs of 22 and 28 for a 22 minute 48 second source video

The model selector appears only at import. For this 22:48 source, Model v1 was quoted 22 credits and Model v2 28 — roughly a 30% premium.

On a 60-credit free month that works out to roughly 60 minutes of source video on v1, or about 46 minutes on v2.

Vizard AI upload screen with Model v2 selected and the Upload button showing a cost of 13 credits for a 10 minute 1 second 1080p video

Model v2 selected on a 10:01 1080p upload: 13 credits, against 10 for the same file on v1.

Watch the billing order. Vizard charges the base per-minute rate the moment you import a video — before you have configured anything or generated a single clip. The v2 premium is then charged separately when clips are generated. If you paste a link, change your mind at the settings screen and walk away, those base credits are already gone.

The Result: 10 Clips vs 5

Same video, same settings, both models:

  Model v1 Model v2
Clips returned 10 5
Average clip length ~38 seconds ~59 seconds
Clip construction One continuous run of the source Several segments spliced together
Scoring Single virality score Virality plus Hook / Flow / Insight / Spread
Titles Descriptive Editorial — capitals, emoji
Export resolution 720p 720p
Watermark Yes Yes

Vizard's own description — v1 is "faster, more clips", v2 gives "fewer clips, more complete" — turned out to be accurate. Exactly half the clips, noticeably longer.

Vizard AI results page for Model v1 showing 10 generated clips in a vertical 9:16 format with virality scores and a visible Vizard watermark

Model v1 returned 10 clips from the 10-minute source, each a single continuous segment, each scored out of 10.

The difference nobody mentions: v2 splices

This is the behavioural change that is not in Vizard's changelog and is worth knowing before you pick.

Every v1 clip was a single continuous run of the source. Start point, end point, done.

v2 assembles clips from multiple non-adjacent moments. Its top clip was built from four separate points in the source — 03:55, then 05:15, 05:23 and 05:54 — stitched into one narrative.

That is a genuine editorial decision, not a longer trim. It means v2 can build a story arc that never existed contiguously in your footage. It also means v2 clips can cut between moments in ways you did not film, so if your video depends on continuity, check v2's output before publishing it.

Vizard AI results page for Model v2 showing 5 clips with virality scores of 10.0 plus Hook, Flow, Insight and Spread sub-scores

Model v2 returned 5 clips from the identical source, with longer runtimes and an extra Hook / Flow / Insight / Spread breakdown under each score.

Longer source, more clips — but not proportionally

We also ran a 22:48 version of the same broadcast through v2. It returned 9 clips — up from 5 on the 10-minute cut, but for 2.3x the source length and 2.8x the credits.

The Scores Are Decorative

v2 shows a headline virality score out of 10 plus four sub-scores. It looks rigorous. We would not lean on it.

The top-scoring clip in our v2 run was rated 10.0 — while the model's own written rationale for that clip said it "lacks a strong hook or high-stakes information, making it the weakest in terms of shareability." Its sub-scores were Hook 7, Flow 7, Insight 8, Spread 9. None of that adds up to a 10.

Another clip scored 9.4 while being described as "somewhat dry" with no "wow moment".

There is also a range problem. Every v2 clip landed between 9.2 and 10.0, so the score cannot separate a strong clip from a weak one. v1's scores at least spread from 7.5 to 9.0.

Read the written rationale, ignore the number. The prose is genuinely useful and often quite honest about a clip's weaknesses. The score is not.

What The Premium Does Not Buy

Two things were identical across both models, and both matter more than clip count for most creators.

Captions. We scored Vizard's transcription against an official reference transcript of the same footage. On conversational speech it managed 2.75% word error rate over a 109-word sample — genuinely good. On technical, jargon-heavy speech it hit 10.0% over a 100-word sample. Every error was a proper noun or domain term: "optical comm" became "optical calm" repeatedly, "phased array" became "phase ray", and astronaut Christina Koch's surname came out as "Cook".

Those exact same errors appeared in both the v1 and v2 output. The clipping model does not touch transcription. If captions are your problem, changing model will not fix it — and note that the free plan shows you the transcript on screen but offers no subtitle file to export.

Vizard AI per-clip export menu on the free plan showing 720p selected and 1080p, 2K and 4K each marked with a paid upgrade badge

The per-clip export menu on the free plan: 720p is selectable, while 1080p, 2K and 4K each carry a paid badge — regardless of which model produced the clip.

Export quality. Both models exported at 720x1280 from a 1920x1080 source, both with the "made with Vizard.ai" watermark burned into the video file. On the free plan, 1080p, 2K and 4K all sit behind the paywall regardless of which model you choose.

The Trap: Your Model Choice Is Account-Wide

This one cost us credits, so learn it for free.

The model dropdown appears only on the import screen. Once a video is imported, open it later and there is no model selector anywhere — which makes it look like the model is locked to that project.

It is not. The model is an account-level setting applied at generation time. We imported a video at the v1 price, later switched the account to v2 for an unrelated test, then came back and generated the first video — and it processed on v2 and charged the premium.

Before you generate an old draft, check which model your account is currently set to. Otherwise you will be charged a premium you did not choose for a video you imported days earlier.

Which Should You Use?

Use Model v1 if you are on the free plan and want the most output per credit, you are testing whether Vizard suits your footage at all, your source is already tightly edited, or you plan to trim clips yourself anyway. Ten usable starting points beat five polished ones when you are the editor.

Use Model v2 if you want clips closer to publishable without manual work, your source is loose and conversational so a spliced narrative genuinely helps, or you are producing a small number of pieces and quality matters more than volume. Just budget for roughly 30% fewer minutes per month.

For most small creators evaluating Vizard on the free plan, start with v1. It costs less per minute, gives you twice as many candidates to judge the tool on, and produces clips whose relationship to your original footage is easy to verify. Move to v2 once you know the tool works for your content and you would rather have five near-finished clips than ten rough ones.

Testing Vizard for the first time? Start on the free plan with Model v1 — it gives you twice as many clips per credit to judge whether the tool suits your footage before you commit to anything.

Try Vizard AI free

FAQ

Is Model v2 available on the free plan?

Yes. Both v1 and v2 are selectable on the free plan with visible prices. v2 is not locked behind an upgrade — it simply costs more credits per minute.

How many credits does Model v2 use?

About 1.3 credits per minute of source video, against roughly 1.0 for v1. A 10-minute video cost 13 credits on v2 and 10 on v1 in our testing.

Does Model v2 give better captions?

No. Caption accuracy was identical in our tests — the same transcription errors appeared in both v1 and v2 output. The clipping model does not affect transcription.

Does Model v2 export at higher quality?

No. Both models exported at 720p on the free plan from a 1080p source, both watermarked. Export resolution is a plan limit, not a model feature.

Can I switch an existing project from v1 to v2?

There is no per-project model selector after import. The model is an account-level setting applied when clips are generated, so changing your account setting and then generating an older draft will process it on the new model — and charge the difference.

Why did Model v2 give me fewer clips?

That is the intended behaviour. Vizard describes v2 as returning "fewer clips, more complete". In our test it returned exactly half as many as v1 from the same source, with each clip about 1.8x longer.