Video Lab model directory

29 frontier video models.One place to direct them.

Video Lab gives creators one prompt-and-control workflow for Veo, Seedance, Kling, Wan, PixVerse, LTX, and more—without juggling provider accounts or unfamiliar APIs.

Editorial selection · Last reviewed 3 October 2026

The screening slate

Choose by the shot you need.

There is no universally best model. Start with the production constraint that matters most—audio, resolution, duration, format, or advanced control—then test the creative direction.

Hover or focus a card to preview motion when available. Every example links to its original source.

Google logo

Google

Original model creator

Veo 3.1

Cinematic campaign shots, product worlds, and scenes where motion and audio should be planned together.

Duration
4, 6, or 8 seconds
Resolution
720p, 1080p, or 4K
Audio
Optional generated audio
Explore model details
Google logo

Google

Original model creator

Veo 3.1 Fast

Faster Veo drafts, campaign tests, and 720p–4K shots where turnaround and cost matter as much as look.

Duration
4, 6, or 8 seconds
Resolution
720p, 1080p, or 4K
Audio
Optional generated audio
Explore model details
Google logo

Google

Original model creator

Veo 3.1 Lite

Lower-cost Veo drafts, 720p tests, and short cinematic clips that do not need 4K or references.

Duration
4, 6, or 8 seconds
Resolution
720p or 1080p
Audio
Optional generated audio
Explore model details
ByteDance logo

ByteDance

Original model creator

Seedance 2.5

Long-form ad concepts, cinematic one-takes, reference-led brand work, and scenes where sound drives action.

Duration
4–30 seconds
Resolution
480p, 720p, or 1080p
Audio
Optional synchronized audio
Explore model details
ByteDance logo

ByteDance

Original model creator

Seedance 2.0

Multi-format ad concepts, longer social shots, and creative teams testing several delivery formats.

Duration
4–15 seconds
Resolution
480p, 720p, 1080p, or 4K
Audio
Optional generated audio
Explore model details
Kuaishou logo

Kuaishou

Original model creator

Kling 3 Pro

Movement-led product shots, character action, and sequences that benefit from explicit shot direction.

Duration
3–15 seconds
Resolution
1080p or native 4K
Audio
Optional generated audio
Explore model details
Kuaishou logo

Kuaishou

Original model creator

Kling 3 4K

Native 4K delivery, movement-led product shots, and sequences that need Kling 3 subject elements at 4K.

Duration
3–15 seconds
Resolution
Native 4K
Audio
Optional generated audio
Explore model details
Kuaishou logo

Kuaishou

Original model creator

Kling 3 Turbo Pro

Faster Kling 3 drafts, native-audio social clips, and 1080p shots that start from text or one still.

Duration
3–15 seconds
Resolution
1080p
Audio
Native audio
Explore model details
Kuaishou logo

Kuaishou

Original model creator

Kling O3

Realistic motion at a lower Kling rate, first/last-frame direction, and multi-shot social clips.

Duration
3–15 seconds
Resolution
Model-managed high-quality output
Audio
Optional generated audio
Explore model details
Alibaba logo

Alibaba

Original model creator

Happy Horse 1.1

Social campaigns that need several aspect ratios, native audio, and straightforward high-resolution delivery.

Duration
3–15 seconds
Resolution
720p or 1080p
Audio
Native audio
Explore model details
Google logo

Google

Original model creator

Gemini Omni Flash

Rapid creative exploration, social hooks, and short audio-visual concepts that still need 1080p or 4K.

Duration
3–10 seconds
Resolution
360p, 720p, 1080p, or 4K
Audio
Native audio
Explore model details
Alibaba logo

Alibaba

Original model creator

Wan 3.0

Longer 1080p campaign and social shots that need image direction and optional sound in one pass.

Duration
2–30 seconds
Resolution
480p, 720p, or 1080p
Audio
Optional native audio
Explore model details
Alibaba logo

Alibaba

Original model creator

Wan 3.0 Prime

Faster long-form campaign and social shots that still need image direction and optional sound.

Duration
2–30 seconds
Resolution
480p, 720p, or 1080p
Audio
Optional native audio
Explore model details
Alibaba logo

Alibaba

Original model creator

Wan 2.7

Repeatable social footage, product cutaways, and teams that value straightforward control.

Duration
2–15 seconds
Resolution
720p or 1080p
Audio
Native generated audio
Explore model details
MiniMax logo

MiniMax

Original model creator

MiniMax H3 Max

Native-audio drafts, social clips, and image-led shots. Generation time varies with the model settings and provider queue.

Duration
5–15 seconds
Resolution
480p, 768p, or 1080p
Audio
Native audio
Explore model details
MiniMax logo

MiniMax

Original model creator

MiniMax H3 Max Turbo

Faster MiniMax drafts, native-audio social clips, and image-led shots that need first/last frames without references.

Duration
5–15 seconds
Resolution
480p, 768p, or 1080p
Audio
Native audio
Explore model details
MiniMax logo

MiniMax

Original model creator

MiniMax H3

High-detail hero shots, atmospheric sequences, and polished brand footage.

Duration
5–15 seconds
Resolution
480p, 768p, 2K, or 4K
Audio
Native audio
Explore model details
xAI logo

xAI

Original model creator

Grok Imagine Video 1.5

Draft-to-final iteration, native-audio social ideas, and teams comparing quality against cost.

Duration
1–15 seconds
Resolution
480p, 720p, or 1080p
Audio
Native audio
Explore model details
xAI logo

xAI

Original model creator

Grok Imagine Video 1.5 Lite

Low-cost Grok drafts, quick social ideas, and native-audio tests before a final pass.

Duration
1–15 seconds
Resolution
480p, 720p, or 1080p
Audio
Native audio
Explore model details
Black Forest Labs logo

Black Forest Labs

Original model creator

FLUX 3

Natural motion from text or a product still, plus restyling an existing clip when you already have motion you want to keep.

Duration
5–20 seconds; Edit clip uses the source MP4 (up to 15s)
Resolution
720p or 1080p; Edit clip is 720p
Audio
Optional generated audio; Edit clip keeps the source clip’s audio
Explore model details
PixVerse logo

PixVerse

Original model creator

PixVerse C1

Film-grade PixVerse shots, identity-led references, and first/last transitions with optional audio.

Duration
1–15 seconds
Resolution
360p, 540p, 720p, or 1080p
Audio
Optional generated audio
Explore model details
PixVerse logo

PixVerse

Original model creator

PixVerse V6

Style exploration, multi-clip concepts, social effects, and rapid draft-to-final workflows.

Duration
1–15 seconds
Resolution
360p, 540p, 720p, or 1080p
Audio
Optional generated audio
Explore model details
Lightricks logo

Lightricks

Original model creator

LTX 2.3 Pro

High-resolution brand footage, motion-conscious production, and shots destined for larger displays.

Duration
6, 8, or 10 seconds
Resolution
1080p, 1440p, or 4K
Audio
Optional generated audio
Explore model details
Kuaishou logo

Kuaishou

Original model creator

Kling O3 Pro

Realistic Kling O3 motion at the Pro rate, first/last-frame direction, and multi-shot social clips.

Duration
3–15 seconds
Resolution
Model-managed high-quality output
Audio
Optional generated audio
Explore model details
Kuaishou logo

Kuaishou

Original model creator

Kling O3 4K

Native 4K O3 motion, first/last-frame direction, and multi-shot social clips that need 4K delivery.

Duration
3–15 seconds
Resolution
Native 4K
Audio
Optional generated audio
Explore model details
Lightricks logo

Lightricks

Original model creator

LTX 2.5 Pro

Newer LTX motion at 1080p, motion-conscious production, and shots that need 24, 25, or 50 fps.

Duration
6, 8, or 10 seconds
Resolution
720p or 1080p
Audio
Optional generated audio
Explore model details
Lightricks logo

Lightricks

Original model creator

LTX 2.5 Fast

Faster LTX drafts, 4K LTX 2.5 delivery, and shots that need 24, 25, 48, or 50 fps.

Duration
6, 8, or 10 seconds
Resolution
720p, 1080p, 1440p, or 4K
Audio
Optional generated audio
Explore model details
OpenAI logo

OpenAI

Original model creator

Sora 2

OpenAI motion, native-audio social clips, and 720p shots from text or a still.

Duration
4, 8, 12, 16, or 20 seconds
Resolution
720p
Audio
Native audio
Explore model details
OpenAI logo

OpenAI

Original model creator

Sora 2 Pro

True 1080p OpenAI motion, native-audio campaign clips, and premium Sora shots.

Duration
4, 8, 12, 16, or 20 seconds
Resolution
720p or true 1080p
Audio
Native audio
Explore model details

Head-to-head

Compare two models, then try both.

Spec tables come from the live Video Lab catalog. No invented scores or reviews.

Wan 3.0 Prime vs Wan 3.0

Same 2–30 second envelope, image modes, and optional native audio. Prime is the accelerated SKU at a higher per-second rate.

Open compare

MiniMax H3 Max vs MiniMax H3

Same 5–15 second native-audio envelope and image modes. H3 Max tops out at 1080p; MiniMax H3 adds 2K and 4K.

Open compare

MiniMax H3 Max vs Seedance 2.5

H3 Max is a 5–15s native-audio 1080p MiniMax SKU. Seedance 2.5 stretches to 30 seconds with optional audio and up to 30 references.

Open compare

Wan 3.0 Prime vs Seedance 2.5

Both reach 30 seconds at up to 1080p with optional audio and first/last frames. Seedance 2.5 takes up to 30 references; Wan 3.0 Prime takes up to ten and adds seed control.

Open compare

Seedance 2.5 vs Kling 3 Turbo Pro

Seedance 2.5 is a long-form reference model. Kling 3 Turbo Pro is a faster 1080p Kling path from text or one still.

Open compare

Veo 3.1 vs Sora 2

Veo 3.1 is a short cinematic Google SKU up to 4K. Sora 2 is OpenAI’s 720p native-audio model with longer duration steps.

Open compare

Veo 3.1 vs Seedance 2.5

Veo 3.1 is a short 4K cinematic SKU. Seedance 2.5 is a longer 1080p native-audio model with a large reference set.

Open compare

Kling 3 Turbo Pro vs Wan 3.0 Prime

Two accelerated SKUs: Kling Turbo Pro is 1080p animate-only. Wan 3.0 Prime is a longer Wan envelope with first/last frames and references.

Open compare

Grok Imagine 1.5 vs Seedance 2.5

Grok is a 1–15s native-audio draft-to-1080p model. Seedance 2.5 is longer, with first/last frames and a much larger reference set.

Open compare

FLUX 3 vs PixVerse V6

FLUX 3 adds Edit clip and 5–20 second shots. PixVerse V6 is a 1–15s style toolkit without Edit clip or multi-image references.

Open compare

LTX 2.3 Pro vs Wan 2.7

LTX 2.3 Pro is a 6/8/10s 4K production SKU with fps control. Wan 2.7 is a 2–15s 1080p generalist with references and native audio.

Open compare

Gemini Omni Flash vs Veo 3.1

Two Google SKUs: Gemini Omni Flash is a fast 3–10s native-audio model. Veo 3.1 is a short cinematic SKU with optional audio and fewer references.

Open compare

Happy Horse 1.1 vs MiniMax H3

Happy Horse 1.1 is a 1080p social-format native-audio model. MiniMax H3 adds first/last frames and 2K/4K at a 5-second minimum.

Open compare

Veo 3.1 Lite vs Veo 3.1 Fast

Lite is the lightest Veo 3.1 SKU at 720p/1080p. Fast adds 4K and up to three references on the same 4–8 second envelope.

Open compare

Kling 3 4K vs Kling 3 Pro

Kling 3 4K is the dedicated native-4K SKU. Kling 3 Pro lists 1080p or native 4K and routes 4K jobs to the same fal 4K endpoints.

Open compare

Kling O3 4K vs Kling O3 Pro

O3 4K is native 4K only. O3 Pro is the higher-rate O3 SKU without a 4K output option. Workflows otherwise match.

Open compare

LTX 2.5 Fast vs LTX 2.5 Pro

Fast is the quicker LTX 2.5 SKU with 1440p, 4K, and 48 fps. Pro stays 720p/1080p with 24, 25, or 50 fps.

Open compare

Grok Imagine 1.5 Lite vs Grok Imagine 1.5

Same 1–15 second native-audio envelope at 480p–1080p. Lite is the lower-rate tier; Grok Imagine 1.5 adds up to seven reference images.

Open compare

One production path

Change the lens, not the workflow.

01

Pick for the constraint

Compare duration, format, audio, resolution, and controls before choosing a model.

02

Direct one clear shot

Write the prompt once, set supported controls, and generate.

03

Keep the result

Every successful clip lands in Video Lab history and your existing Clip Studio Library.

Ready before you generate

The quote changes with the shot.

Model, duration, resolution, and audio all matter.

Video Lab quotes each setup against this month’s plan before you generate. If a generation fails without producing usable output, that usage is returned to the monthly limit.

Questions, answered

Video Lab FAQ

What is Clip Studio Video Lab?

Video Lab is one workspace for generating videos with text prompts, source images, or an existing clip across 29 curated frontier models. Choose a model, add a prompt, image, or source MP4, generate, and keep the completed clip in your Clip Studio Library.

Are these the 29 most popular AI video models?

No objective cross-provider popularity ranking exists. This is an editorial selection of 29 frontier video models chosen for useful differences in quality, speed, control, format support, audio, and resolution.

How is Video Lab priced?

Each plan includes a monthly generation limit. Model, duration, resolution, and other settings use that allowance. Upgrade anytime for a higher monthly limit.

Can I compare the same prompt across models?

Yes. Generate with one model at a time, then use Try another model to preserve your prompt, compatible settings, and available source images while you select a different model.

Which image-to-video workflows are supported?

Every Video Lab model can animate a starting image. Veo, Seedance, Kling 3 Pro, Kling 3 4K, Kling O3, Kling O3 Pro, Kling O3 4K, Wan, MiniMax, Gemini, PixVerse, LTX, and FLUX 3 also support first-to-last-frame direction, including MiniMax H3 Max Turbo. Kling 3 Turbo Pro is animate-only—no first/last frames. Veo 3.1 and Veo 3.1 Fast, Seedance, Happy Horse, Gemini, Wan, MiniMax H3, MiniMax H3 Max, Kling O3, Kling O3 Pro, Kling O3 4K, PixVerse C1 (up to 4), and Grok offer multi-image reference workflows. Veo 3.1 Lite is text, animate, and first/last only—no references or 4K. Kling 3 Pro and Kling 3 4K use subject reference elements inside the image workflow. H3 Max Turbo has first/last like H3 Max but no references. Sora 2 and Sora 2 Pro animate a starting image; they do not take first-to-last frames or references. FLUX 3 can also restyle an existing MP4 with Edit clip (up to 15 seconds and 50 MB).

Can Video Lab edit an existing clip?

Yes, on FLUX 3. Choose Edit clip, pick a library clip or upload an MP4. Clips must be MP4, under 15 seconds and 50 MB. Lip sync follows new dialogue, including translations. Duration and aspect follow the source. Output is 720p and keeps the source clip’s audio unless you rewrite or translate the dialogue. Other Video Lab models generate from text or images.

The prompt is ready

Now choose its lens.

Start with one model. Keep the prompt. Try another. Every successful result stays in your Clip Studio Library.

Open Video Lab