Vital signs
Why it's on the table
On the table, Video (Vi) is seat 20 of 58, in the Creative & Design family. It is an emerging element — the job is real and here to stay, but the leaderboard still changes quarterly. Choose for this quarter, hold loosely, and watch the changelog. It is optional: plenty of companies run without it — until a specific trigger (scale, regulation, cost, or customers) makes it essential for them. It sits in the mid price band — a real line item that should earn its keep visibly.
Video: the top 5 — v2026.Q3
Edition v2026.Q3 · ranking, pricing and status verified 2026-08-06.
1Veo (3.1 / Flow)Google DeepMind
Google AI Pro $19.99/mo (1,000 Flow credits) · Ultra $99.99–199.99/mo (10k–25k credits) · API ~$0.10–0.40/s 720p–1080p, up to $0.60/s 4K+audio (via fal, Aug 2026)Best for The default: best blend of quality, native audio, control features, and sheer availability — in Gemini, Flow, the API, and inside rivals' own products.
Veo 3.1 generates 1080p/4K with native dialogue, SFX and ambient audio, camera controls, character consistency, and scene extension to ~148 seconds — and it ships everywhere, including inside Runway's own paid tiers. Google also holds the #1 spot on the Artificial Analysis text-to-video arena outright with Gemini Omni Flash (1,244 Elo, May 2026), so the quality crown stays in-house even as Veo 3.1 itself ages.
Watch Veo 3.1 has slid to mid-table on the crowd arena (#11, 1,093 Elo) as Chinese models passed it — Google's answer (Omni Flash) lives in Gemini, not the Veo API, muddying the lineup. Base generations are still 8s clips, and 4K-with-audio at $0.60/s adds up fast.
2Seedance 2.5ByteDance Seed
Dreamina free daily credits · API usage-based via BytePlus/fal/Replicate (2.0 Pro/Fast/Mini tiers) · unlimited via partners (e.g. Higgsfield, from Aug 7, 2026)Best for Actual storytelling — 30-second single generations, native multi-shot, and multi-reference control (up to 9 images, 3 clips, 3 audio files) instead of stitched 8-second fragments.
Seedance 2.0 holds #1 on the image-to-video arena (1,197 Elo) and #3 on text-to-video (1,224, Mar 2026), with 1.3M+ Replicate runs; 2.5 extends single generations to 30 seconds (twice extendable) with green-screen editing, camera blocking, and reference-video interpretation. No other model turns a script into a coherent multi-shot sequence with this little glue work.
Watch The IP overhang is real: Disney cease-and-desists over celebrity/character output forced ByteDance to pause the 2.0 global launch on Mar 15, 2026, and first-party pricing remains opaque — access flows through Dreamina, aggregators, and partner resellers. TikTok-adjacent geopolitics is a standing platform risk.
#1 image-to-video (1,197 Elo) and #3 text-to-video (1,224) on the arena (Aug 2026; model dated Mar 2026) [src] · Seedance 2.5: 30s single generation, 2x extendable, reference-video and green-screen control (vendor page, Aug 2026) [src] · Global 2.0 launch paused Mar 15, 2026 over Disney/studio IP demands after Feb 2026 China launch [src]3Runway (Gen-4.5)Runway AI
Free 125 one-time credits · Standard $12/mo (annual; $15 monthly, 625 credits) · Pro $28/35 (2,250) · Max $76/95 (9,500) · Enterprise customBest for Founders producing finished pieces, not clips — generation plus editing (Aleph), performance transfer (Act-Two), upscaling, and now other vendors' models in one workflow.
Gen-4.5 (Dec 2025 update) does native audio and minute-long multi-shot generation with character consistency, and the surrounding toolchain — Aleph edit-anything, Topaz 4K upscale, custom voices — is what turns raw generations into watchable video. The company raised $315M at $5.3B (Feb 10, 2026) and hedged model risk by hosting Veo 3.1 and others, adding a model router in July 2026.
Watch Runway doesn't enter the crowd arena, and its 'world's best video model' claim dates to the Nov 2025 launch — on raw generation quality the Chinese labs have likely passed it. Credit economics bite at scale, and the strategic center of gravity is drifting to world models (GWM-1) over creator video.
4Kling 3.0Kuaishou
API ~$0.112/s · $0.168/s with audio · $0.196/s with voice control (via fal, Aug 2026); consumer tiers on kling.ai (page JS-rendered — see notes)Best for The best price-to-cinema ratio — 15-second 1080p clips with native audio and voice control at roughly a dollar per shot.
Kling 3.0 (Feb 2026) sits top-10 on the text-to-video arena in three variants (Pro 1080p at 1,111 Elo) and generates up to 15s with native audio, lip-sync, and a voice-control tier — at $0.112–0.196/s, roughly half Veo's standard rate. For volume experimentation it is the strongest cost-quality point among the closed leaders.
Watch First-party global pricing is hard to even read (JS-walled site), documentation trails the US vendors, and Kling is one product inside a short-video giant whose priorities are its own feed. Elo sits a tier below Seedance/Omni Flash/H3.
5Hailuo (MiniMax H3)MiniMax
Consumer subs via hailuoai.video (page JS-rendered — see notes) · API usage-based via MiniMax platform / fal / Replicate, standard + fast tiersBest for Budget wildcard with crowd-validated quality — the #2-rated model overall, native 2K output, and cheap fast variants for iteration.
H3 (Jul 2026) rates 1,238 Elo — #2 on text-to-video, within 6 points of Google's leader — and #3 on image-to-video, generating native 2K from text or images with omni-style multimodal context. MiniMax is now a public company (HKEX IPO Jan 9, 2026), taking the going-dark risk off the table.
Watch 2025 revenue was just $79M against frontier-scale compute burn — the subsidized pricing may not hold. The ecosystem is thin (no editing suite, patchy docs in English), and neither the consumer nor first-party API price list could be verified from vendor pages.
Video: the top 8 compared
Edition v2026.Q3 · ranking, pricing and status verified 2026-08-06.
| Tool | Base clip | Max res | Native audio | Multi-shot / consistency | Open weights | API $/s (typ.) | Arena Elo T2V (Aug 2026) |
|---|---|---|---|---|---|---|---|
| Veo 3.1 | 8s (ext. ~148s) | 4K | Yes | Scene extension · character consistency | No | $0.10–0.60 | 1,093 (#11) |
| Seedance 2.5 | 30s (2x extend) | 1080p+ | Yes | Native multi-shot · 9-img/3-clip/3-audio refs | No | unpublished | 1,224 (2.0, #3) |
| Runway Gen-4.5 | up to 60s | 1080p (Topaz 4K upscale) | Yes (Dec 2025) | Character-consistent shots · Aleph editing | No | credit-based | not entered |
| Kling 3.0 | 15s | 1080p | Yes + voice control | Omni variant · lip-sync | No | $0.11–0.20 | 1,111 (#7) |
| Hailuo H3 | n/p | 2K | Yes (omni context) | Multimodal single-context | No | n/p (std + fast tiers) | 1,238 (#2) |
| Wan 2.7 | n/p | 1080p | n/p | S2V · Animate (2.2 line) | ≤2.2 only (Apache-2.0) | cheapest via aggregators | 1,161 (#4) |
| LTX-2.3 | 20s | 1080×1920 | Yes (joint 14B+5B) | — | Yes (free <$10M rev) | usage API or self-host | not listed |
| Gemini Omni Flash | n/p | n/p | Yes (omni) | In-Gemini generation | No | Gemini API | 1,244 (#1) |
How to choose your video
- If you can only hold one subscription and need reliable quality with audio, everywhere
- Veo via Google AI Pro ($19.99/mo, 1,000 Flow credits) — accept 8s base clips and extend; move to the API when volume justifies $0.10–0.40/s.
- If the job is a story — ads, trailers, episodic shorts — not isolated clips
- Seedance 2.5: 30s single generations with multi-reference control beat stitching 8s fragments, if you accept the IP-friction availability risk and reseller-mediated pricing.
- If you edit as much as you generate, or deliver client-ready pieces
- Runway — Gen-4.5's minute-long consistent multi-shot plus Aleph/Act-Two/upscaling is the only real post-production stack, and its model router hedges the generation layer.
- If cost per watchable second decides, and you iterate in volume
- Kling 3.0 (~$0.11/s, audio $0.168/s) or Hailuo fast tiers — the arena says you give up little quality for roughly half Veo's rate.
- If you need open weights — on-prem, fine-tuning, or no per-second meter
- LTX-2.3 (22B open weights, 20s joint audio-video, free under $10M revenue) or the Wan 2.2 Apache lineage; accept a quality tier below the closed leaders.
Video: the whole field
18 more tools tracked in this category, including 5 dead, renamed, or sunsetting — a reference that hides the graveyard isn't one. Verified 2026-08-06.
| Tool | Maker | What it is | Entry | Status |
|---|---|---|---|---|
| Gemini Omni Flash | Native video generation inside Gemini — #1 on the text-to-video arena (1,244 Elo, May 2026); effectively Veo's sibling and successor-in-waiting | Google AI plans from $4.99/mo | active | |
| Wan | Alibaba | Wan 2.7 is arena #4 (1,161 Elo, Apr 2026); open-weight lineage ends at Wan 2.2 (Apache-2.0, 16.9k stars) — newer versions are API-first; cheapest volume option on aggregators | free weights (2.2) · API tokens | active |
| LTX-2.3 | Lightricks | 22B open-weight dual-stream model: 20s joint audio-video at 1080×1920, free for companies under $10M revenue; API, desktop app, or self-host | free <$10M rev · usage API | active |
| Luma Dream Machine (Ray3.14) | Luma AI | Ray3.14/Ray3.2 with reframe and video-to-video; clean credit pricing but absent from arena top ranks | Plus $30/mo · Pro $90 · Ultra $300 | active |
| Pika 2.5 | Pika Labs | Early mover now a social-effects niche player; still shipping but off the quality leaderboards | free · $8–76/mo | fading |
| Vidu Q3 | Shengshu | Start-end-to-video up to 16s (Q3 Pro/Turbo); strong in anime/reference-to-video | free credits · subs | active |
| PixVerse v5.6 | AIsphere | Unit-priced volume play popular for social content | free tier · unit pricing | active |
| Grok Imagine Video 1.5 | xAI (listed as SpaceXAI on arena) | ~30s generation latency, #4 on image-to-video arena (1,113 Elo, May 2026); bundled with X/Grok subs | via X Premium / SuperGrok | active |
| SkyReels V4 | Skywork AI | Arena top-10 both directions (1,107 Elo T2V, Mar 2026); film-oriented | free credits · subs | active |
| HappyHorse 1.1 | Alibaba-ATH | Alibaba's second video lab — #5 T2V (1,148 Elo, Jun 2026); two Alibaba entries now sit in the arena top 8 | n/p | active |
| MAGI-2 Preview | Sand.ai | Autoregressive architecture, arena debut Aug 2026 (#6 I2V) — the one to watch for long-form | preview | active |
| Moonvalley Marey | Moonvalley | Fully-licensed training data, 'commercially safe' positioning for studios; Realism v1.5, on fal/ComfyUI since Aug 2025 | app subs · API | active |
| Higgsfield | Higgsfield AI | Pivoted from own models to an aggregator suite — resells Seedance 2.0/2.5 (unlimited tier from Aug 7, 2026) and MiniMax H3 with viral presets | subs (frequent promos) | active |
| Hunyuan Video | Tencent | 2024's open-weight pioneer; quiet through 2026 as Wan and LTX took the open mantle | free weights | fading |
| Adobe Firefly Video | Adobe | Commercially-safe generation inside Premiere/Firefly app; enterprise indemnification is the pitch, quality mid-tier (single-source; not independently benchmarked here) | Firefly plans · CC bundle | active |
| Mochi 1 | Genmo | Open-weight 2024 entrant; development quiet through 2026 | free weights | fading |
| Stable Video Diffusion | Stability AI | The category's first open weights (2023); lineage abandoned amid Stability's restructuring | — | dead |
| Sora / Sora 2 | OpenAI | The graveyard headliner: app discontinued Apr 26, 2026 (announced Mar 24), API shuts Sep 24, 2026. Peaked ~3.3M downloads/mo (Nov 2025), collapsed under 500k users at ~$1M/day burn; Disney's $1B partnership got under an hour's notice | — | dead |
Video: the category in numbers
Edition v2026.Q3 · ranking, pricing and status verified 2026-08-06.
- OpenAI killed Sora (announced Mar 24, 2026; app off Apr 26, API Sep 24): users collapsed ~1M → <500k, ~$1M/day compute burn, $2.1M lifetime IAP revenue — proof that raw generation without a workflow is a money pit [src]
- Runway raised $315M at $5.3B (Feb 10, 2026), then launched a model router (Jul 23, 2026) — the workflow layer is aggregating the generation layer [src]
- MiniMax IPO'd on HKEX Jan 9, 2026 (2025 revenue $79M) — first pure-play model lab in this category to go public [src]
- Chinese labs hold 8 of the arena's top 10 text-to-video slots (Aug 2026): ByteDance, MiniMax, Alibaba x2, Kuaishou, Skywork — only Google represents the US in the top tier [src]
- IP enforcement now gates releases: Disney cease-and-desists ('virtual smash-and-grab') paused Seedance 2.0's global launch Mar 15, 2026; expect provenance and likeness controls as table stakes [src]
- Market sizing is contested: Grand View pegs AI video generation at just $946M for 2026 (20.3% CAGR to $3.4B by 2033) — conservative against vendor raises and compute burn; treat all category TAM figures skeptically [src]
Video: method & sources
Ranking charter: editorial, weighing crowd-arena Elo, controllability/workflow depth, availability, and price — no affiliate consideration. Biggest conflict resolved: the Artificial Analysis arena would order this Gemini Omni Flash > MiniMax H3 > Seedance > Wan, but we rank the products founders can actually build a pipeline on — Omni Flash is folded into the Veo/Google entry (same maker, in-Gemini access), and Runway ranks on workflow despite not entering the arena at all. Replicate still describes Gen-4.5 as 'ranked #1 on the Artificial Analysis benchmark' — that claim dates to its Nov 2025 launch and is stale; we cite the live Aug 2026 leaderboard instead. Pricing weaknesses flagged: Kling and Hailuo consumer tiers could not be read from their JS-rendered vendor pages (API prices cited from fal, an aggregator); Seedance has no public first-party price list — access is via Dreamina credits and resellers. Sora's exact shutdown dates (app Apr 26, API Sep 24, 2026) come from OpenAI's notice as relayed at commission time; TechCrunch confirmed the shutdown and its causes but not those two dates — re-verify before quoting. Wan's openness ends at 2.2 (Apache); the arena-ranked 2.7 is API-only. Grand View's market number is single-source and looks understated. Adjacent elements: talking-head/presenter tools (HeyGen, Synthesia) → Av · Avatar Presenter; image models incl. Midjourney's video experiments → Im · Image; voice/dubbing → Vo · Voice; app-builders → Vb. Ranking criteria: verified commercial traction, independent satisfaction surveys, agent benchmarks, and founder-fit (price floor, lock-in, surfaces). Editorial, never paid — the charter. Machine-readable twin: vi.json.
All sources (22)
- https://artificialanalysis.ai/video/leaderboard/text-to-video
- https://artificialanalysis.ai/video/leaderboard/image-to-video
- https://deepmind.google/models/veo/
- https://gemini.google/subscriptions/
- https://fal.ai/models/fal-ai/veo3.1
- https://fal.ai/models/fal-ai/kling-video/v3/pro/text-to-video
- https://seed.bytedance.com/en/seedance2_5
- https://runwayml.com/pricing
- https://techcrunch.com/2026/02/10/ai-video-startup-runway-raises-315m-at-5-3b-valuation-eyes-more-capable-world-models/
- https://techcrunch.com/2025/12/11/runway-releases-its-first-world-model-adds-native-audio-to-latest-video-model/
- https://techcrunch.com/2026/03/24/openais-sora-was-the-creepiest-app-on-your-phone-now-its-shutting-down/
- https://techcrunch.com/2026/03/29/why-openai-really-shut-down-sora/
- https://techcrunch.com/2026/03/15/bytedance-reportedly-pauses-global-launch-of-its-seedance-2-0-video-generator/
- https://replicate.com/collections/text-to-video
- https://en.wikipedia.org/wiki/MiniMax_(company)
- https://ltx.io/model
- https://lumalabs.ai/pricing
- https://www.pika.art/pricing
- https://github.com/Wan-Video/Wan2.2
- https://higgsfield.ai/
- https://www.moonvalley.com/
- https://www.grandviewresearch.com/industry-analysis/ai-video-generator-market-report
Our take
Improving faster than any other creative element. Budget experiments, not commitments.
Combines with
This is element 20 of 58. The table is versioned quarterly — when a tool loses its seat, the changelog records the succession.
Explore the full table →