VidModelHub

Updated 2026-08-15 · Checked daily, updated at least weekly

Best AI video models right now

The flagship video generators, ranked by what they're actually best at — with real outputs, blind-vote arena scores and side-by-side comparisons. Rumors labeled, prices sourced, nothing quoted from memory.

Editor's pick

Veo 3.1

Best all-around today — but blind-vote arenas are a real fight: Seedance 2.0 and Kling 3.0 trade the top spot depending on the test.

Best video model for each job

Text-to-video Arena (no audio)

Source: Artificial Analysis · as of Jul 2026

1
Gemini Omni FlashGoogle
1327
2
HappyHorse-1.0Alibabalimited access
1287
3
HappyHorse-1.1Alibabalimited access
1274
4
Seedance 2.0 720pByteDance
1272
5
Kling 3.0 1080p (Pro)Kuaishou
1245

Elo scores are volatile and some leaders (HappyHorse, Seedance 2.0) have limited public access. This is the no-audio board; Gemini Omni Flash also tops the separate with-audio track. Neither FLUX 3 (gated early access since Jul 23) nor MiniMax H3 (launched Jul 31) has appeared here yet, so launch-day ranking claims for both are unverified.

What the newest models actually produce

Official launch clips from the three models that shipped in nine days — sound on.

Seedance 2.5 — 30 seconds, one take30s

The longest single generation any flagship offers, unedited and with sound.

MiniMax H3 — native 2K with stereo15s · 2K

From MiniMax's own API docs. The only one of the three you can buy today.

FLUX 3 — 15s alpine descent15s

Black Forest Labs' longest published demo; the model itself is still application-gated.

Official launch clips — each vendor's own demo, sourced on its model page below · mirrored and re-encoded here, sound intact.

Prompt and output, paired

Hover to play — every clip ships with the exact prompt that made it (Seedance 2.0, from ByteDance's prompt guide).

Seedance 2.0Marketing

Product ad: golden horse pendant

Reference the running form of the horse from Video 1, generate a golden horse galloping on a grassland, then freeze-frame its magnificent running pose, transforming into a horse-shaped gold pendant.

Uses reference inputs
Seedance 2.0Character

Consistent character, new scene

Reference the woman's appearance from Image 1, Image 2, and Image 3, generate a scene of her eating cake at a coffee shop.

Uses reference inputs
Seedance 2.0Film

Choreographed fight scene

Reference the character actions and camera language from Video 1, generate a fight scene between Image 2 and Image 1. Image 2 is the character on the left, Image 1 is the character on the right. With intense background music.

Uses reference inputs
Seedance 2.0VFX

Golden particle effects transfer

Reference the golden particle effects from Video 1, have the character from Image 2 play a flute while surrounded by the same particle effects.

Uses reference inputs
Seedance 2.0Architecture

FPV dive concept video

Reference the camera movement from Video 1 to create a concept video for a tech park. Use the high-rise building from Image 1 as the visual center, with the same first-person diving perspective, highlighting the tech aesthetic of the park.

Uses reference inputs
Seedance 2.0Editing

Scene-to-scene particle transition

Video 1, at the moment the leaf touches the ground, golden particle effects burst out, a gust of wind blows, then connect to Video 2.

Uses reference inputs

Demo outputs and prompts from ByteDance's official Seedance 2.0 prompt guide. Hover to play. Our own same-prompt tests replace these at each model launch.

Same prompt, every model

Head-to-head breakdowns on the specs that decide a shot — not a beauty contest.

Video tools & guides

Get video-model drop alerts

One email when a major video model ships or gets a confirmed date. No spam, unsubscribe anytime.

Full release tracker

Upcoming & rumored

AnnouncedJul 23, 2026 · early access
FLUX 3Black Forest Labs

One network for image, video with native audio (up to 20s) and robot action prediction. Video is open to approved applicants only; there is no public API, no published pricing and no open weights yet.

AnnouncedAug 2026 (est.) · “within days”

MiniMax says it will publish H3's weights “within the next few days, subject to the relevant laws and regulations”, and designed the model for compatibility with several domestic Chinese accelerators. No repo has appeared on the MiniMaxAI Hugging Face org yet. Would be the strongest openly licensed video model to date.

Announced“Soon” after Jul 31 launch

ByteDance says the Volcano Ark API is coming shortly after the Jul 31 consumer launch. Until it lands there is no per-second price, no third-party endpoint and no access outside Jimeng and Doubao Pro. The three insider-sourced dates before launch (Jul 6, 13, 16) all failed, so treat any new date without an official post the same way.

RetiringSep 24, 2026

Two-stage wind-down of Sora 2: the consumer app and sora.com already went offline on Apr 26, 2026; the developer API (sora-2, sora-2-pro) is decommissioned Sep 24, 2026. Teams on Sora endpoints need a migration path before this date.

RumoredH2 2026 (est.)

Not announced. Estimate based on Kuaishou's historical 4–6 month release cadence since Kling 3.0 (Feb 2026).

AnnouncedLater in 2026
FLUX 3 Dev (open weights)Black Forest Labs

Confirmed but undated: the open-weight release of the full multimodal backbone — the first openly licensed model generating video, audio and images from one set of weights. No size or license terms announced; FLUX.1/FLUX.2 Dev tiers shipped non-commercial.

RumoredUnknown
Veo 4Google DeepMind

Not announced. Google I/O 2026 shipped Gemini-family updates, not Veo 4. Rumors point to native 4K and longer clips.

RumoredNot announced
Sora 3OpenAI

No Sora 3 has been announced. After shutting down Sora 2, OpenAI said it is repurposing Sora as a world-model research effort — there is no successor product or date on the roadmap.

Shipped

ReleasedJul 31, 2026
MiniMax H3MiniMax

Hailuo's third generation, launched as a general-purpose omni model: one network reads text, images, video and audio as a single context and returns 2K video with native stereo sound, 5–15s. Live the day it was announced — in the Hailuo app, MiniMax's own API and third-party resellers. Open weights promised within days.

ReleasedJul 31, 2026
Seedance 2.5ByteDance

Released Jul 31, 2026 after a Jun 23 preview. 30-second single generations with multi-round extension into minutes, up to 30 images + 10 videos + 10 audio as references in one request, and timestamp-addressed editing (green screen, viewpoint, camera move). Live in Jimeng AI and Doubao Pro; the Volcano Ark API is “coming soon” and no 2.5 model ID exists in the ModelArk catalog yet. No international access, no pricing.

ReleasedFeb 12, 2026
Seedance 2.0ByteDance

Launched Feb 12, 2026 with Doubao 2.0. Multi-modal input (up to 9 images, 3 videos, 3 audio files), native audio, 480p–4K output. ByteDance paused its global rollout in mid-March 2026 after cease-and-desist letters from Disney, Paramount, Warner Bros. and the MPA over copyrighted characters in outputs.

ReleasedFeb 4, 2026
Kling 3.0Kuaishou

Multi-shot storyboard generation — one prompt produces a sequence of coherent shots.

ReleasedOct 2025
Veo 3.1Google DeepMind

Google's current flagship video model. Strong physics and native audio; premium API pricing.