Skip to content
NewNew models in the catalogue
Musevate
Model library

Every model that matters, one workflow

Musevate routes each generation to the strongest compatible model for your mode, format and quality tier, so you get the best result without learning any of them. Every clip below was generated on Musevate; models we have not filmed yet show their name instead.

75 of 75

Wan 2.2

~2 min

The dependable all-rounder for everyday production work.

T2VI2V720p10s

from 20 credits / 5s · about $0.90

Wan 3.0

Q8 · ~3 min

Smoother motion and cleaner scenes than Wan 2.2, up to 1080p.

T2VI2VRefs1080p10s

from 25 credits / 5s · about $1.12

Kling v3

Q8 · ~6 min

Cinematic framing with fluid, believable movement.

T2VI2V4K15s

from 40 credits / 5s · about $1.80

Gemini Omni Flash 1.1

Q8 · ~45 s

Google's Gemini Omni Flash 1.1, with sound.

T2VI2VRefs4K10s

from 25 credits / 5s · about $1.12

Wan 3.0 Prime

Q10 · ~2 min

The flagship Wan tier, for shots that carry the piece.

T2VI2V1080p30s

from 35 credits / 5s · about $1.57

MiniMax H3 Max

Q8 · ~10 s

H3 retrained for closer prompt adherence.

T2VI2VRefs1080p10s

from 14 credits / 5s at 480p · about $0.63

Seedance 2.5

Q9 · ~2 min

Single shot clips up to 30 seconds, with audio.

T2VI2VRefs1080p30s

from 115 credits / 5s · about $5.16

Grok Imagine 1.5 Lite

~45 s

Grok Imagine 1.5 for a fraction of the price.

T2VI2V1080p15s

from 10 credits / 5s · about $0.45

Vidu Q4

~3 min

Image-led video with sound, up to 4K.

I2VRefs4K16s

from 25 credits / 5s · about $1.12

Kandinsky 6.0 Pro

~14 min

Kandinsky Lab's flagship, with sound.

T2VI2V480p5s

from 70 credits / 5s at 480p · about $3.14

LTX 2.3 Video Fast

The previous LTX generation, tuned for speed.

T2VI2V4K20s

from 23 credits / 5s at 1080p · about $1.03

Veo 3.1

~60 s

The current flagship, strongest on realism and natural motion.

T2VI2V4K8s

from 100 credits / 5s · about $4.49

Kandinsky 6.0 Lite

~3 min

Kandinsky's fast model: five seconds, with sound.

T2VI2V480p5s

from 12 credits / 5s at 480p · about $0.54

Wan 2.7

~4 min

Wan 2.7, with first and last frame control.

T2VI2VRefsLip sync1080p15s

from 25 credits / 5s · about $1.12

Pixelcut Looping Video

~15 s

A product photo, turned into a seamless loop.

I2V1080p15s

from 20 credits / 5s at 480p · about $0.90

Veo 3.1 Fast

~35 s

Veo 3.1 accelerated, for iterating on a shot quickly.

T2VI2V4K8s

from 40 credits / 5s · about $1.80

Kling v2.1 Master

~4 min

The premium 2.1 tier, with stronger dynamics.

T2VI2V1080p10s

from 70 credits / 5s at 1080p · about $3.14

Kling v2.6

~2 min

Kling 2.6 Pro, cinematic image-to-video with fluid motion.

T2VI2V1080p10s

from 35 credits / 5s at 1080p · about $1.57

Kling v2.0

~5 min

Kuaishou's Kling 2.0, in 5 and 10 second clips.

T2VI2V720p10s

from 70 credits / 5s · about $3.14

Hailuo 2.3

~85 s

High fidelity model tuned for realistic human motion.

T2VI2V1080p10s

from 15 credits / 5s at 768p · about $0.67

Wan 2.1 1.3b

~20 s

The small, cheap Wan for quick 480p drafts.

T2V480p5s

from 10 credits / 5s at 480p · about $0.45

Pixverse v4.5

Q6 · ~60 s

PixVerse v4.5 refines v4 with smoother motion, better handling of complex actions and closer prompt following. It makes 5 or 8 second clips at up to 1080p.

T2VI2V1080p8s

from 30 credits / 5s · about $1.35

Hailuo 02

~3 min

Hailuo 2, strong on physical realism.

T2VI2V1080p10s

from 11 credits / 5s at 768p · about $0.49

Sora 2 Pro

~3 min

The most capable Sora tier, with sound.

T2V720p12s

from 120 credits / 5s · about $5.39

P-Video-2 Pro

~10 s

Pruna's top model, run in its quality mode.

T2VI2V768p15s

from 13 credits / 5s at 480p · about $0.58

Grok Imagine Video

~45 s

xAI's Grok Imagine video model.

T2VI2V720p15s

from 15 credits / 5s · about $0.67

Sora 2

~2 min

Flagship generation with synchronised sound.

T2V720p12s

from 25 credits / 5s · about $1.12

Veo 3 Fast

~35 s

Veo 3 at a lower price and a shorter wait.

T2VI2V1080p8s

from 40 credits / 5s · about $1.80

Kling v2.5

~3 min

Kling 2.5 Turbo Pro, professional output at speed.

T2VI2V1080p10s

from 20 credits / 5s at 1080p · about $0.90

Seedance 2.0 Fast

~90 s

Seedance 2.0 at speed.

T2VI2VRefs720p15s

from 45 credits / 5s · about $2.02

Pixverse v5

~85 s

Create 5s-8s videos with enhanced character movement, visual effects, and exclusive 1080p-8s support. Optimized for anime characters and complex actions

T2VI2V1080p8s

from 30 credits / 5s · about $1.35

Video 01

~4 min

The original Hailuo model, in 6 second clips.

T2VI2V720p5s

from 25 credits / 5s · about $1.12

Ray 2 720p

~85 s

Luma Ray 2 at 720p, in 5 and 9 second clips.

T2VI2V720p9s

from 45 credits / 5s · about $2.02

Seedance 2.0 Mini

~2 min

The smallest Seedance 2.0 tier.

T2VI2VRefs720p15s

from 30 credits / 5s · about $1.35

Seedance 1.5 Pro

~75 s

Joint audio and video, following complex instruction.

T2VI2V1080p12s

from 15 credits / 5s · about $0.67

Video 01 Director

~2 min

Video-01 with explicit camera direction.

T2VI2V720p5s

from 25 credits / 5s · about $1.12

Pixverse v4

Q6 · ~55 s

PixVerse v4 is the fourth generation of PixVerse's video model. It makes 5 or 8 second clips from a prompt or a still image, at resolutions up to 1080p.

T2VI2V1080p8s

from 30 credits / 5s · about $1.35

Ray Flash 2

~30 s

Ray 2's look, faster and cheaper, at 720p.

T2VI2V720p9s

from 15 credits / 5s · about $0.67

P-Video-2

~15 s

Pruna's quality-focused P-Video, up to 1080p.

T2VI2VLip sync1080p20s

from 10 credits / 5s · about $0.45

P-Video

Q4 · ~15 s

Fast iteration, with a draft mode.

T2VI2VLip sync1080p20s

from 5 credits / 5s · about $0.22

Gen 4.5

~2 min

Runway Gen-4.5, strong motion quality and prompt adherence.

T2VI2V720p10s

from 30 credits / 5s · about $1.35

Veo 2

~2 min

Google Veo 2, faithful to written direction.

T2VI2V720p8s

from 120 credits / 5s · about $5.39

LTX 2.5 Fast

~25 s

Built for speed, with native audio and high resolutions.

T2VI2V4K20s

from 12 credits / 5s · about $0.54

MiniMax H3 Max Turbo

MiniMax H3 Max Turbo

S9

H3 Max with the brakes off, at half the cost.

T2VI2V1080p15s

from 8 credits / 5s at 480p · about $0.36

MiniMax H3

MiniMax H3

~6 min

MiniMax H3: video with sound from a prompt, a frame or references.

T2VI2VRefs4K10s

from 14 credits / 5s at 480p · about $0.63

FLUX 3 Video

FLUX 3 Video

Q8 · ~70 s

Coherent, physically plausible motion from a single frame.

T2VI2V1080p20s

from 45 credits / 5s · about $2.02

Kling v2.5 Standard

Kling v2.5 Standard

S7

The Kling 2.5 look, at the standard tier.

I2V720p10s

from 10 credits / 5s · about $0.45

Happy Horse 1.1

Happy Horse 1.1

Happy Horse 1.1, refined over the first release.

T2VI2V1080p15s

from 35 credits / 5s · about $1.57

Gemini Omni Flash

Gemini Omni Flash

Google's Gemini Omni Flash, the release before 1.1.

T2VI2VRefs720p10s

from 35 credits / 5s · about $1.57

Kling AI Avatar v2 Standard

Kling AI Avatar v2 Standard

Turns a portrait into a speaking avatar.

Lip sync720p10s

from 15 credits / 5s · about $0.67

Bytedance Omnihuman v1.5

Bytedance Omnihuman v1.5

Animates a human subject from a single image.

Lip sync1080p10s

from 40 credits / 5s · about $1.80

Seedance 1.0 Pro

Seedance 1.0 Pro

Seedance 1 Pro for text-to-video.

T2VI2V1080p12s

from 15 credits / 5s · about $0.67

PixVerse v6

PixVerse v6

Pixverse v6, strong on stylised and animated looks.

T2VI2V1080p15s

from 15 credits / 5s · about $0.67

Kling Video v3 Turbo Pro

Kling Video v3 Turbo Pro

The professional Kling v3 tier, accelerated.

T2VI2V1080p15s

from 35 credits / 5s at 1080p · about $1.57

Veo 3.1 Lite

Veo 3.1 Lite

The lightest Veo 3.1 tier.

T2VI2V1080p8s

from 14 credits / 5s · about $0.63

Kling 2.1

Kling 2.1

Kling 2.1 image-to-video, a proven older generation.

I2V720p10s

from 15 credits / 5s · about $0.67

Wan 2.5

Wan 2.5

Preview build of Wan 2.5.

T2VI2VLip sync1080p10s

from 25 credits / 5s · about $1.12

Kling Video v3 Standard Turbo

Kling Video v3 Standard Turbo

~2 min

Kling v3 quality at turbo speed.

T2VI2V720p15s

from 30 credits / 5s · about $1.35

Seedance v1 Pro Fast

Seedance v1 Pro Fast

~50 s

Seedance 1 Pro text-to-video, accelerated.

T2VI2V1080p12s

from 6 credits / 5s · about $0.27

Kling O3

Kling O3

~14 min

The Kling O3 standard tier.

T2VI2VRefs4K15s

from 30 credits / 5s · about $1.35

Happy Horse

Happy Horse

Alibaba's Happy Horse text-to-video model.

T2VI2V1080p15s

from 35 credits / 5s · about $1.57

Kling 1.6

Kling 1.6

Kling 1.6 Standard, Kuaishou's release of December 2024.

T2VI2V720p10s

from 15 credits / 5s · about $0.67

Animate Diff MotionLoRA

Animate Diff MotionLoRA

Q3 · ~4 min

This AnimateDiff build adds MotionLoRAs for camera moves (zoom, pan, tilt, roll), animating Stable Diffusion image models into short clips from a prompt.

T2V480p5s

from 7 credits / 5s at 480p · about $0.31

Animate Diff

Animate Diff

Q3 · ~4 min

AnimateDiff, an open research model, adds a motion module to Stable Diffusion image models so they can animate. It makes short 512x512 clips from a prompt.

T2V480p5s

from 8 credits / 5s at 480p · about $0.36

Wan 2.1

Wan 2.1

~3 min

Accelerated Wan 2.1 image-to-video at 720p.

T2VI2V720p5s

from 19 credits / 5s at 480p · about $0.85

Ltx Video

Ltx Video

Q5 · ~15 s

LTX-Video is the first DiT-based video generation model capable of generating high-quality videos in real-time. It produces 24 FPS videos at a 768x512 resolution faster than they can be watched.

T2VI2V480p10s

from 5 credits / 5s at 480p · about $0.22

Grok Imagine 1.5

Grok Imagine 1.5

~35 s

Expressive motion from a still, with sound.

T2VI2V1080p15s

from 20 credits / 5s · about $0.90

Video 01 Live

Video 01 Live

~3 min

Trained for Live2D and illustrated characters.

I2V720p5s

from 25 credits / 5s · about $1.12

Hailuo 2.3 Fast

Hailuo 2.3 Fast

~70 s

Hailuo 2.3 with lower latency, keeping its motion quality.

I2V1080p10s

from 12 credits / 5s at 768p · about $0.54

Text2video Zero

Text2video Zero

Q2 · ~2 min

Text2Video-Zero, from Picsart AI Research, makes a text-to-image diffusion model generate video with no training on video. It makes short 512x512 clips.

T2V480p5s

from 10 credits / 5s at 480p · about $0.45

Pia

Pia

Q3 · ~55 s

PIA, the Personalized Image Animator from OpenMMLab, is an open research model shown at CVPR 2024. It animates a still image to follow a short text prompt.

I2V480p5s

from 5 credits / 5s at 480p · about $0.22

Seedance 1 Lite

Seedance 1 Lite

~40 s

The light Seedance tier, for text and image driven clips.

T2VI2V1080p12s

from 10 credits / 5s · about $0.45

Ltx Video 0.9.7 Distilled

Ltx Video 0.9.7 Distilled

Q5 · ~50 s

LTX-Video 0.9.7 Distilled is Lightricks' distilled LTX-Video 13B: faster, at a slight cost in quality. It works from a prompt or a still image at 30fps.

T2VI2V720p16s

from 10 credits / 5s · about $0.45

Veo 3

Veo 3

~85 s

Google's Veo 3, with generated audio.

T2VI2V1080p8s

from 100 credits / 5s · about $4.49

Seedance 2.0

Seedance 2.0

~2 min

ByteDance's multimodal generation with native audio.

T2VI2VRefs4K15s

from 55 credits / 5s · about $2.47

Narrow it down

Ninety-odd models is a lot to read through. These are the same catalogue, sorted by the questions people actually ask of it.

Head to head: Sora 2 vs Veo 3.1 and Kling 3 vs Veo 3.1.

Choosing a model

How many AI video models can I use on Musevate?

Musevate gives access to 75 AI video models through one account, drawn from providers including Google, Alibaba, Kuaishou, ByteDance, MiniMax, Lightricks, Luma, OpenAI and xAI. New models are added as they are released, and your workflow does not change when they are.

Which is the cheapest AI video model?

The lowest priced model on Musevate, Ltx Video, starts at 5 credits for a 5 second clip at 480p. The Fast tier does not choose it on its own: Fast routes to the cheapest model that can do your request among those rated at least 6 out of 10 for quality that make a 5 second clip within 3 minutes, which starts at 6 credits for a 5 second clip at 720p. You can pick Ltx Video by hand under Advanced.

What is the difference between text-to-video and image-to-video?

Text-to-video generates a clip from a written description alone. Image-to-video animates a still image you upload, using your prompt to direct the motion while keeping the subject recognisable. Most models on Musevate support one or both, and each model page states which.

Do I have to pick a model myself?

No. Choose Fast, Quality or Cinema and Musevate routes the request to the strongest compatible model. Picking a specific model is available under Advanced for people who want it.

Pick a tier, not a model

100 credits for $4.49 is about 16 clips on Fast, or 4 on a typical model, and the router chooses the right model for every shot.

Create your first video
AI Video Models · Musevate