MarketingDecision Matrix12 min readPublished July 24, 2026

Six contenders · longest single pass 30s · cheapest leader $6/min · no model wins every axis

FLUX 3 vs Seedance, Omni & Kling: Mapping the Video Field

FLUX 3 arrived July 23 with 20-second clips and zero public pricing. Seedance 2.5’s API opened July 16 with unstitched 30-second passes. Gemini Omni Flash holds all four Artificial Analysis Video Arena boards at $6.00 per minute. Six models, four axes — duration, price, control, and audio — and no single winner.

DA
Digital Applied Team
Senior strategists · Published Jul 24, 2026
PublishedJul 24, 2026
Read time12 min
SourcesBFL · AA · TechTimes
Longest single pass
30s
Seedance 2.5 · API GA Jul 16
FLUX 3 max clip
20s
native audio · gated early access
Arena #1 price
$6/min
Gemini Omni Flash · AA leaderboard
FLUX 3 vs Seedance 2.0
52%
vendor preference — a coin flip

The best AI video model in July 2026 depends entirely on which axis you measure. FLUX 3 landed on July 23 with the longest single-pass clip Black Forest Labs has ever published — 20 seconds with native audio — but no pricing, no independent benchmark, and gated access. Seedance 2.5 and Gemini Omni Flash, meanwhile, are shipping in production today.

The stakes are practical, not academic. Marketing and creative teams are budgeting real money against per-minute API rates that range from $6.00 to $24.00 on the Artificial Analysis leaderboard, while two of the most-hyped models in the field — FLUX 3 and Seedance 2.5 — publish no official per-minute price at all. Vendor comparison charts all favor their own model, and almost none put duration, price, control surface, and legal risk on the same page.

This guide does exactly that. We map six models — FLUX 3 Video, Seedance 2.5, Gemini Omni Flash, HappyHorse 1.1, Kling 3.0, and Veo 3.1 — across every axis a buying team actually weighs, mark every unverifiable cell as “not published” rather than guessing, and close with a which-model-for-which-job decision matrix.

Key takeaways
  1. 01
    FLUX 3 is the wildcard, not the leader.Its 20-second single-pass claim is the longest BFL has published, but there is no public pricing, no Artificial Analysis Video Arena entry, and BFL’s own preference chart is labeled a preliminary eval of an early candidate — a coin-flip 52% against Seedance 2.0 and Gemini Omni Flash.
  2. 02
    Seedance 2.5 owns duration and control.A native, unstitched 30-second clip in one API call — the longest single-pass commercial generation as of the July 16 API opening — plus up to 50 multimodal reference inputs and region-level editing. No official per-minute price has been published.
  3. 03
    Gemini Omni Flash owns rank and price.It leads all four Artificial Analysis Video Arena boards (Text-to-Video with audio Elo 1,245) and is the cheapest top-tier model at $6.00 per minute — but clips cap at 10 seconds at 720p/24fps.
  4. 04
    Copyright risk belongs in the decision matrix.Seedance carries live legal exposure: cease-and-desist letters from five major studios, the MPA’s first-ever AI cease-and-desist, and the Andersen v. Stability AI trial scheduled for September 8, 2026 that could set precedent for the whole category.
  5. 05
    No single model wins every axis.Duration leader (Seedance 2.5), price and leaderboard leader (Omni Flash), premium 4K option (Veo 3.1), and the unpriced wildcard (FLUX 3) trade off differently — route by workload, not by headline.

01The FieldOne week that redrew the video field.

Two launch events frame this comparison. On July 16, ByteDance opened public API access to Seedance 2.5 via BytePlus — completing the rollout announced at Volcano Engine’s FORCE conference on June 23. Seven days later, on July 23, Black Forest Labs announced FLUX 3, its first multimodal frontier model, jointly trained across image, video, and audio in one architecture. Between those two dates sits a field that already looked crowded: the post-Sora AI video market has been consolidating fast since OpenAI discontinued the Sora consumer app (announced March 24, offline April 26), with the Sora API itself sunsetting September 24, 2026.

That leaves six models worth a marketing team’s attention today — and three distinct storylines among them.

The wildcard
FLUX 3 Video
20s single pass · native audio · gated

Longest single-pass claim BFL has published, but no pricing, no independent benchmark, and the vendor’s own eval is labeled preliminary. Promising — unverifiable as of July 24.

Early Access · bfl.ai
The production leader
Seedance 2.5
30s single pass · 50 reference inputs

The duration and control-surface leader, API-live since July 16 via BytePlus. Carries unresolved copyright exposure that belongs in any client-facing risk assessment.

API GA · BytePlus
The value leader
Gemini Omni Flash
10s clips · $6.00/min · Arena #1

Ranked first on all four Artificial Analysis Video Arena boards and the cheapest top-tier option. Short clips, but unbeatable for iterative concept work.

GA · artificialanalysis.ai #1

02Black Forest LabsFLUX 3: the wildcard with the least public data.

FLUX 3 is Black Forest Labs’ first multimodal frontier model — jointly trained across image, video, and audio rather than bolting a video head onto an image backbone. The headline capability: up to 20 seconds of video with native audio in a single generation, the longest single-pass duration BFL has published. Preliminary evaluations were run at 720p on 10-second clips. For the full architecture story, launch tiers, and the robotics angle, see our full FLUX 3 launch breakdown.

What shipped is narrower than the announcement suggests. FLUX 3 Video is in gated Early Access — API plus private weights by request. FLUX 3 Image is “coming in weeks” and had not shipped as of July 24. FLUX 3 Action is limited to selected robotics partners, and FLUX 3 Dev — the open-weight multimodal backbone — is a future release with no date. Martin Scorsese reportedly joined BFL as an advisor in early July, a signal of how seriously the lab is courting film-adjacent credibility. Early testers named at launch include Canva, Burda, Magnific, Krea, and Picsart.

No price exists
As of July 24, 2026, no public pricing has been announced for any FLUX 3 tier. VentureBeat notes enterprise buyers cannot yet calculate total cost of ownership. Any per-minute or per-clip figure you see attached to FLUX 3 right now is a guess — and FLUX 3 has no Artificial Analysis Video Arena entry either, so there is no independent quality benchmark to price against.

The only comparative data available is BFL’s own preference testing — and BFL itself labels the chart a “preliminary evaluation of an early FLUX 3 candidate,” meaning the tested checkpoint is not necessarily the model now in Early Access. Read the numbers with that caveat attached: the blowout wins come against the field’s trailing models, while against the two models that actually lead the arena — Seedance 2.0 and Gemini Omni Flash — FLUX 3 scores 52%, a statistical coin flip.

FLUX 3 preference-test win rate · BFL’s own evaluation

Source: BFL launch blog, Jul 23, 2026 — vendor-stated, preliminary eval of an early FLUX 3 candidate
vs Luma Ray 3.2FLUX 3 preferred in 93% of comparisons
93%
vs Runway Gen-4.5FLUX 3 preferred in 77%
77%
vs Grok Imagine VideoFLUX 3 preferred in 69%
69%
vs Kling v3 ProFLUX 3 preferred in 60%
60%
vs Happy Horse v1FLUX 3 preferred in 59%
59%
vs Happy Horse 1.1FLUX 3 preferred in 57%
57%
vs Seedance 2.052% — a statistical coin flip
52%
vs Gemini Omni Flash52% — a statistical coin flip
52%
“You can’t cheat reality. A model that only learns images can only generate images. But the world is not made of still frames. It moves, sounds, changes, and responds.”— Robin Rombach, co-founder and CEO, Black Forest Labs

03ByteDanceSeedance 2.5: the duration and control leader.

Seedance 2.5 generates a native, unstitched 30-second clip in a single API call — the longest single-pass commercial video generation available as of the July 16 API opening. For context from TechTimes’ read of the competitive field: Veo 3.1 caps native passes at 8 seconds and Runway Gen-4.5 at 5–10 seconds. We covered Seedance 2.5’s 30-second single-pass launch in depth when the rollout was announced; the API opening makes it real for production teams.

The control surface is the other differentiator. Seedance 2.5 accepts up to 50 multimodal reference inputs — images, video clips, and audio — in one generation request, up from 15 total in Seedance 2.0 (9 images, 3 video, 3 audio). That is more than a threefold jump, and it matters most for brand work: product shots, brand footage, and voice references can all steer one generation. Region-level editing is new too — modifying a specific face, object, or background element in an already-generated clip without regenerating the whole shot.

Under the hood, ByteDance describes a Sparse Diffusion Transformer built by its Doubao team, using sparse attention across the full clip duration rather than a sliding window to hold character, lighting, and camera-position consistency over 30 seconds, with audio co-generated in the same latent space as the video rather than post-synced. Treat those as vendor-stated specs: no independent third-party benchmark of Seedance 2.5 existed as of the API opening, and the model has no confirmed Artificial Analysis leaderboard entry — the arena still lists Seedance 2.0.

Single-pass duration
Unstitched, one API call
30s

The longest single-pass commercial generation as of the Jul 16 API opening — no stitching, no extension passes, per TechTimes’ competitive review.

Veo 3.1 native pass: 8s
Reference inputs
Multimodal control slots
50

Images, video clips, and audio combined in one request — up from 15 total in Seedance 2.0. The widest control surface in this comparison set.

Seedance 2.0: 15
Enterprise ARR
Vendor-stated, Jun 23
$2B

ByteDance cited $2 billion in annualized recurring revenue for the Seedance enterprise platform at the FORCE conference — before the public API even opened.

Vendor figure — unaudited

One number Seedance 2.5 does not have: a price. ByteDance has published no official per-minute API rate for Seedance 2.5 as of this writing. The Artificial Analysis leaderboard entry at $9.07 per minute is for Seedance 2.0 720p — a different model — and third-party reseller estimates for 2.5 are not ByteDance’s published rate. In our comparison table below, that cell reads “not published,” because it is.

04GoogleGemini Omni Flash: the leaderboard and price leader.

If Seedance owns duration, Gemini Omni Flash owns everything measurable. As of July 24, it is ranked #1 on all four Artificial Analysis Video Arena boards: Text-to-Video with audio (Elo 1,245), Text-to-Video without audio (1,326), Image-to-Video with audio (1,201), and Image-to-Video without audio (1,375). It is also the cheapest top-tier model on the leaderboard at $6.00 per minute for 1080p API generation — $0.10 per second, or roughly $1.00 for a 10-second 720p clip per VentureBeat’s pricing table. Since Gemini Omni’s video-generation debut, Google has merged generation and conversational editing into one interface, which is where the workflow advantage compounds.

The constraints are equally concrete. Clips run 10 seconds maximum (3-second minimum) at 720p/24fps — a 30-second social cut means three generations and a stitch. And editing of uploaded (rather than model-generated) video is unavailable in the EEA, Switzerland, and the UK — a real limitation for European teams hoping to iterate on existing brand footage.

API price per minute of generated video · AA-standardized

Source: Artificial Analysis Video Arena, Text-to-Video with audio — Elo, ranks, and standardized API pricing retrieved Jul 24, 2026
Gemini Omni FlashAA rank #1 · Elo 1,245 (T2V w/ audio)
$6.00/min
Seedance 2.0 720pAA rank #2 · Elo 1,227 — 2.5 price not published
$9.07/min
HappyHorse 1.1AA rank #4 · Elo 1,149
$9.90/min
Kling 3.0 1080p (Pro)AA rank #6 · Elo 1,111
$20.16/min
Veo 3.1AA rank #11 · Elo 1,095 · native 4K
$24.00/min

Read the price chart alongside the ranks and the field inverts the usual assumption that quality costs more: the #1-ranked model is the cheapest, and the most expensive model in the set ranks eleventh. Veo 3.1’s premium buys native 4K and per-second billing, not arena preference. For high-volume iterative work — concept testing, storyboard animatics, ad-variant drafts — Omni Flash at $6.00 per minute is four times cheaper than Veo 3.1 ($24.00 ÷ $6.00 = 4×) and undercuts every AA-listed rival in this set.

05The Chasing PackHappyHorse, Kling, and Veo: the rest of the field.

Alibaba HappyHorse 1.1 (June 23, 2026) ranks #4 on the AA Text-to-Video-with-audio board (Elo 1,149) and #4 on Image-to-Video with audio (1,110), at $9.90 per minute. It generates 15-second clips at 1080p/24fps with native audio — lip-sync plus Foley. Its predecessor, HappyHorse 1.0 (April 2026), still ranks #5 (Elo 1,128) at $13.20 per minute, and reportedly took the #1 arena spot within days of its April 7 debut, beating both Seedance 2.0 and Kling 3.0 at the time. Alibaba’s Wan2.7 line holds #3 (Elo 1,164) on the same board — we unpack Alibaba’s two-line video stack in a companion post. The oft-repeated claim that HappyHorse was built under Zhang Di, the engineer behind Kling 1.0 and 2.0, comes from aggregator coverage rather than Alibaba itself — treat it as reported, not confirmed.

Kling 3.0 (Kuaishou, released February 2026 — not a new launch this window) sits mid-pack: the Pro 1080p tier ranks #6 (Elo 1,111, $20.16/min), Standard 720p #9 (1,100, $15.12/min), with the Omni variants at #10 and #12. That makes Kling the most expensive per-minute line in this comparison except Veo 3.1. Its control-surface gap is stark: per TechTimes’ competitive read, Kling 3.0 accepts text and image inputs only — no video or audio reference input — against Seedance 2.5’s 50-slot multimodal system.

Google Veo 3.1 remains the production Veo line — no new Veo number shipped in this window. It ranks #11 (Elo 1,095) at $24.00 per minute, the most expensive model in the set, but offers native 4K and per-second billing, which is why it keeps its place in premium broadcast workflows. Veo 3.1 Lite (#13, Elo 1,092, $4.80/min) and Veo 3.1 Fast (#14, Elo 1,091) fill out the budget tiers — with a pricing wrinkle on Fast we address under the table below.

06Decision TableThe video field at a glance — July 2026.

No published source puts all six models on one table with duration, audio, control, arena rank, price, and availability together — vendor pages show only their own model, and the AA leaderboard has rank and price but not duration or control detail. This table is the union. Every cell traces to a cited source; where no official figure exists, the cell says so rather than guessing.

Comparison of FLUX 3 Video, Seedance 2.5, Gemini Omni Flash, HappyHorse 1.1, Kling 3.0 Pro, and Veo 3.1 across maximum single-pass duration, native audio, reference inputs, Artificial Analysis Video Arena rank and Elo, API price per minute, and availability, as of July 24, 2026.
ModelMax single passNative audioReference inputsAA Arena (T2V w/ audio)API price / minAvailability
AA = Artificial Analysis Video Arena, retrieved Jul 24, 2026 · “not published” = no official figure exists
FLUX 3 Video20s (evals ran at 720p/10s)Yes — nativenot publishednot yet rankednot publishedGated Early Access (API + weights by request)
Seedance 2.530s — unstitchedYes — co-generated in latent space50 multimodal (image / video / audio)not listed (2.0 720p: #2 · 1,227)not publishedAPI GA via BytePlus (Jul 16)
Gemini Omni Flash10s (3s min) @ 720p/24fpsYesConversational editing; uploaded-video edit unavailable in EEA / CH / UK#1 · 1,245$6.00GA
HappyHorse 1.115s @ 1080p/24fpsYes — lip-sync + Foleynot published#4 · 1,149$9.90API live
Kling 3.0 (Pro 1080p)not publishedYes — ranked on AA with-audio boardText + image only#6 · 1,111$20.16API live
Veo 3.18s per pass (extendable)Yes — ranked on AA with-audio boardnot published#11 · 1,095$24.00GA · native 4K, per-second billing

One cell we deliberately left out of the table: Veo 3.1 Fast. Its two sourced prices do not reconcile. Artificial Analysis lists $9.00 per minute; VentureBeat’s per-clip table implies roughly $6.00 per minute ($1.00 per 10-second clip × 6). For Veo 3.1 itself the same cross-check reconciles cleanly ($4.00 per 10 seconds × 6 = $24.00/min, matching AA), so the Fast discrepancy is real, not a rounding artifact. We present both figures with their sources rather than picking one.

The trend the table exposes: the field is racing on different axes because nobody can win them all at once. Long single passes demand attention architectures that hold consistency over 30 seconds; cheap iteration demands short clips and aggressive serving economics; premium broadcast demands 4K pipelines. Each vendor has optimized for the axis its distribution favors — ByteDance for social-native long takes, Google for volume and integration, Kuaishou and Alibaba for the Chinese creator stack. Expect the axes to converge slowly, not suddenly: the physics of serving cost means the 30-second-single-pass and $6-per-minute clubs will stay separate clubs for a while.

Most “which video model” comparisons omit legal exposure entirely. For an agency advising clients, it belongs in the table. Seedance leads on spec but carries the field’s most concrete legal risk: after Seedance 2.0’s February 12, 2026 China launch went viral with unauthorized celebrity likenesses, five major studios — Disney, Warner Bros. Discovery, Paramount Skydance, Netflix, and Sony Pictures — sent cease-and-desist letters, and on February 22 the Motion Picture Association sent its first-ever AI cease-and-desist to a major generative-AI company.

As of July 2026, no federal lawsuit has been filed — serving a Beijing-headquartered company through the Hague Service Convention is estimated at 18–24 months — which means the exposure is unresolved rather than resolved. Industry sentiment captures the ambivalence: Seedance has been described by media advisor Peter Csathy as “the most powerful video generator in the market right now,” even as studio approval remains informal at best.

“Within the industry, I know that a lot of studios haven’t approved Seedance, but yet with a wink and a nod, they’re allowing Seedance to be used. It’s kind of like a ‘don’t ask, don’t tell’ kind of a thing.”— Joel Kuwahara, animation producer, via the Los Angeles Times
Date to watch
The Andersen v. Stability AI trial is scheduled for September 8, 2026 in the Northern District of California — potentially the first judicial ruling on whether AI-generated images and video are infringing derivative works. A plaintiff-friendly ruling would directly sharpen Seedance’s exposure and reshape risk assessments for every model on this page. If your brand work touches recognizable IP, build the contingency in now.

08Decision MatrixWhich model for which job.

Vendor comparisons all favor their own model. A marketing-team lens routes by workload instead — here is how we would assign the field today, with the caveat that every recommendation assumes you run a small paid pilot on your own briefs before committing a quarter’s production budget.

30s single-take social ad
One unstitched pass, no seams

Seedance 2.5 is the only model here that generates a native 30-second clip in one API call. Factor the unresolved copyright exposure into client-facing work — and note no official price is published yet.

Pick Seedance 2.5
Iterative concept work
Cheapest volume iteration

Gemini Omni Flash at $6.00/min is the cheapest top-tier model and #1 on every AA board. Ten-second caps matter less when you’re testing twenty concepts before lunch. EEA/CH/UK teams: uploaded-video editing is unavailable.

Pick Gemini Omni Flash
Brand-asset control
Consistency from reference inputs

Seedance 2.5’s 50-slot multimodal reference system (vs Kling 3.0’s text-and-image-only inputs) is the widest control surface for holding brand look, product detail, and voice across a campaign.

Pick Seedance 2.5
Premium 4K broadcast spot
Native 4K, per-second billing

Veo 3.1 is the most expensive in the set at $24.00/min and ranks #11 on arena preference — but native 4K output and per-second billing keep it the default for broadcast-grade deliverables.

Pick Veo 3.1

And FLUX 3? Until pricing, open access, and an independent benchmark exist, it is a watch-list item, not a production decision. The 20-second single-pass claim with native audio is genuinely differentiated — but a 52% coin flip against the two arena leaders, from a vendor-run preliminary eval of a checkpoint that may not match the shipping model, is not a basis for moving budget. If you want help wiring any of these models into a real pipeline — creative testing on paid channels, always-on social production, or cost modeling per deliverable — our social media team runs exactly these evaluations, and our paid media practice can pressure-test whether AI video actually lifts your creative performance before you scale it.

09ConclusionA field with four leaders and no champion.

The video field, July 2026

Route by axis — duration, price, control, and risk — not by headline.

The July 2026 video field has no single winner because the axes genuinely trade off. Seedance 2.5 owns duration and control — 30-second unstitched passes and 50 reference inputs — with unresolved legal exposure and no published price. Gemini Omni Flash owns rank and price — #1 on all four Artificial Analysis boards at $6.00 per minute — with 10-second clips. Veo 3.1 owns the premium 4K lane at premium cost. And FLUX 3 owns the possibility space: the longest single-pass claim of the week, with nothing independently verifiable behind it yet.

Looking forward, two events could reorder this table before the year ends. FLUX 3’s open-weight Dev backbone — if it ships — would be the first open multimodal video foundation in the field, and open weights have repeatedly reset pricing floors in adjacent categories. And the September 8 Andersen v. Stability AI trial could turn copyright from a soft caveat into a hard procurement gate. Teams that build model-routing flexibility now — rather than standardizing on one vendor — will absorb both shocks cheaply.

The practical move this week is unglamorous: pick one workload, run the same three briefs through Seedance 2.5 and Omni Flash, price the results per finished deliverable, and let your own footage — not anyone’s preference chart — decide.

Put AI video to work in campaigns

The right video model is the one that wins your brief.

Our team helps marketing and creative organizations evaluate AI video models against real briefs — duration, cost per deliverable, brand control, and legal risk — and wire the winners into production workflows.

Free consultationExpert guidanceTailored solutions
What we work on

AI video engagements

  • Model bake-offs on your own briefs — Seedance / Omni / Veo
  • Cost-per-deliverable modeling across per-minute API rates
  • Brand-consistency pipelines with reference-input control
  • Copyright and likeness risk review for client-facing work
  • Always-on social video production at agency scale
FAQ · AI video field guide

The questions we get every week.

FLUX 3 is Black Forest Labs’ first multimodal frontier model, announced July 23, 2026, jointly trained across image, video, and audio in one architecture. Its video tier generates up to 20 seconds with native audio in a single pass — the longest duration BFL has published. Practically, though, access is limited: FLUX 3 Video is in gated Early Access (API plus private weights by request), FLUX 3 Image was still “coming in weeks” as of July 24, FLUX 3 Action is restricted to selected robotics partners, and the open-weight FLUX 3 Dev backbone has no release date. No public pricing has been announced for any tier, so teams cannot yet calculate cost of ownership. Treat it as a watch-list model rather than a production option this quarter.
Related dispatches

Continue exploring AI video.