Video models in 2026: from demo to ad
In one year, OpenAI shut down Sora, Google moved from Veo to Gemini Omni, and models from Alibaba, MiniMax and ByteDance reached the top of the rankings. Here is what already works for a brand, what doesn't yet, what it costs and how to label it in Europe.
As of September 2026, generative video for brands is split between Google's Gemini Omni Flash and Veo 3.1, Kling 3.0, Seedance, Wan 3.0, MiniMax H3, Runway, Luma and FLUX 3, at list prices from $0.05 to $0.80 per second and shots of 3 to 30 seconds. OpenAI has left the market: the Sora API shut down on September 24.
It works for short spots, social content, previs and DOOH loops; not yet for long continuity, exact brand assets or faces without clean clearance. Since August 2, 2026, whoever publishes a deepfake in the EU must disclose it from first exposure.
The verified landscape as of September 2026
Models whose specs and prices we confirmed in each vendor's documentation.
| Model (vendor) | Length per generation | Max resolution | Native audio | Access | USD per second |
|---|---|---|---|---|---|
| Gemini Omni Flash (Google) | 3 to 10 s, extendable to 40 s | 4K (upscaled) | Yes | Gemini API, Google Cloud, Flow | About 0.10 (720p) to 0.30 (4K) |
| Veo 3.1, Fast and Lite (Google) | 4, 6 or 8 s, extendable to 148 s | 4K (Lite: 1080p) | Yes; silent is cheaper on Google Cloud | Gemini API, Google Cloud, Flow | 0.40 (4K: 0.60); Fast 0.10 to 0.30; Lite 0.05 to 0.08 |
| Kling 3.0 (Kuaishou) | 3 to 15 s | 4K | Yes | App and API | 0.084 to 0.168; 4K: 0.42 |
| Seedance 2.0 and 2.5 (ByteDance) | 4 to 15 s (2.0); 4 to 30 s (2.5) | 4K (2.0); 1080p (2.5) | Yes | BytePlus API | 2.0: 0.15 to 0.78; 2.5: 0.23 to 0.57 |
| Wan 3.0 (Alibaba) | Up to 30 s | 1080p | Ranked in with-audio arenas | Model Studio API | 0.05 (480p) to 0.20 (1080p) |
| MiniMax H3 | 4 to 15 s | 2K (regenerated) | Yes, stereo | API and open weights under a community license | About 0.13 |
| Runway Gen-4.5 | Not confirmed | Not confirmed | Yes | App and API | 0.12 |
| Luma Ray3.2 | Up to 20 s | 1080p HDR | Not advertised | App and API | 0.06 (720p) to 0.24 (1080p) |
| FLUX 3 (Black Forest Labs) | Up to 20 s | 4K | Optional, no extra charge | API | 0.17 (HD) to 0.80 (4K) |
| LTX-2.5 (Lightricks) | Not confirmed | 4K (upscaled) | Yes | Open weights; free under $10 million revenue | Your compute |
| HunyuanVideo 1.5 (Tencent) | About 5 s | 720p (1080p with super-resolution) | No | Open weights; license excludes the EU | Your compute |
Google now recommends Omni Flash as the default model and keeps Veo 3.1 for scene extension or last-frame control. In the preference rankings (Artificial Analysis on September 26, Arena on the 21st), the top spots belong to Gemini Omni Flash, Wan 3.0, MiniMax H3 and Seedance 2.x; the 2025 leaders (Kling, Runway Gen-4.5, Veo 3.1, Luma and Sora 2 Pro) now sit mid-table.
The Sora shutdown: a lesson in vendor risk
OpenAI launched Sora 2 on September 30, 2025 and opened it in its API a week later. On March 12, 2026 it expanded that API (20 s clips, 1080p, reusable characters) and twelve days later announced it would retire it. The app and web experience closed on April 26 and the API shut down on September 24, 2026, with no replacement. According to TechCrunch, the app was burning about $1 million a day in compute while its users fell below 500,000, and Disney's $1 billion deal collapsed before any money changed hands.
- Model-agnostic workflow: script, storyboard, key frames and prompts are yours; the video engine should be swappable within days.
- A second engine, tested: validate at least two models on your key frames before committing to a campaign schedule.
Download and archive every approved take the same day, with its prompt and source frame. Google, for example, deletes Veo videos from its servers after two days.
What works in production and what doesn't
Documented cases show the real scale. For the Kalshi ad aired during the 2025 NBA Finals, made mainly with Veo 3, its creator reported 300 to 400 generations for 15 usable clips, one person, two to three days and about $2,000. Coca-Cola used five specialists, more than 70,000 clips and about 30 days for its 2025 holiday ads, and Svedka spent about four months rebuilding a character for its Super Bowl LX spot (February 8, 2026).
Works today
- Short spots and social content cut from shots a few seconds long.
- Previs and animatics: AMC Networks agreed in 2025 to use Runway to pre-visualize shows.
- Variants at scale: according to Google, WPP uses Veo and Imagen to turn one creative idea into hundreds or thousands of versions.
- Loops for screens and backdrops for live experiences.
Not yet
- Long continuity: each generation lasts seconds, Veo extensions drop to 720p and a stable character takes weeks.
- Exact brand assets: logos, packaging and legal lines are composited in post.
- Real faces without clean clearance: vendors block or restrict them, especially in Europe.
- French voice without testing: Kling 3.0 speaks English, Chinese, Japanese, Korean and Spanish, and Seedance 2.5 advertises 11 languages, Spanish among them.
The main risk is reception. Coca-Cola drew criticism again in 2025, and McDonald's Netherlands pulled an AI-made holiday ad within four days (December 6 to 10, 2025). IAB measures the gap: 82% of ad executives believe young consumers like AI ads, against 45% of those consumers.
A production workflow in seven steps
The one we recommend at Slash to go from demo to an asset you can air.
- Script: length, formats (16:9, 9:16), with or without sound, markets, languages and what must be real.
- Storyboard: shots a few seconds long, because no verified model goes beyond 30 s per generation.
- Approved key frames: generated with an image model and product references; the client approves stills, which cost cents, before paying for video.
- Image-to-video: each shot starts from its approved frame, with first and last frame when the model allows it, and several takes per shot.
- Edit: cut, grade and upscale in your editor; logo, packshot and legal lines are composited on top.
- Sound: native audio is fine for drafts; the final mix uses licensed music and contracted voices.
- Compliance: consents, licenses, preserved provenance, market-appropriate labeling and archived masters.
Steps 1 to 3 follow the method in image models in 2026, and we cover sound in voice and audio AI.
What it costs: an illustrative calculation
Price per second matters less than the usable-take rate.
Illustrative assumptions: a 30-second spot with 10 shots; each take is generated at 8 s in 1080p and 1 in 20 is kept, the ratio the Kalshi ad's creator reported. Total: 200 takes and 1,600 generated seconds.
| Model and mode | USD per second | Cost of 1,600 s |
|---|---|---|
| Veo 3.1 Lite, 1080p | 0.08 | $128 |
| Veo 3.1 Fast, 1080p | 0.12 | $192 |
| Gemini Omni Flash, 1080p | About 0.15 | $240 |
| Kling 3.0, 1080p with audio | 0.168 | $268.80 |
| Veo 3.1, 1080p silent (Google Cloud) | 0.20 | $320 |
| Wan 3.0, 1080p | 0.20 | $320 |
| Seedance 2.0, 1080p | 0.37 | $592 |
| Veo 3.1, 1080p with audio | 0.40 | $640 |
The generation bill ranges from about $130 to $640. The multiplier that matters is the usable-take rate: if it halves, the cost doubles. And human work dominates: the Kalshi ad cost about $2,000 in total according to its creator, and McDonald's Netherlands' holiday production kept about 10 people busy for five weeks, according to NBC. For screens without sound, silent Veo 3.1 on Google Cloud costs half.
Faces, voices and consent
Vendors already limit real people, and more so in Europe. In the EU, the UK, Switzerland and the Middle East and North Africa, Veo only generates adults. Gemini Omni Flash does not allow editing uploaded videos in the European Economic Area, Switzerland and the UK, restricts images with minors there and prevents certain recognizable people from appearing in uploaded images.
The law points the same way. In Colombia, Law 2502 of 2025 aggravates impersonation done with AI, and Law 1581 requires authorization to process personal data, a person's image included; in France, influencer advertising with generated faces must carry the mention Images virtuelles. For actors and ambassadors, the contract should name the AI use, media, territories, duration and fee; for the public at an activation, explicit consent and a deletion policy. More on impersonation risk in deepfake fraud.
Article 50: how to label a video in the EU
Since August 2, 2026, whoever publishes a deepfake in the EU must disclose it clearly and at the latest at first exposure (Article 50, paragraphs 4 and 5). The AI Act calls a deepfake content that resembles existing persons, objects, places, entities or events and would falsely appear authentic (Article 3, point 60): a photorealistic video of a real product in a real city can qualify. Providers, for their part, must mark outputs in a machine-readable way; Veo and Omni put SynthID on every clip.
The Code of Practice of June 10, 2026 makes it concrete: an icon or label perceivable at first exposure without user action, for example top right, and for video, at the start and at regular intervals, both online and offline. Evidently artistic or fictional works may carry it in notes or credits, provided it is perceivable at first exposure. On September 24 the Commission published optional icons in SVG and PNG.
It also applies to agencies outside the EU when the video is distributed in the Union (Article 2), with fines of up to €15 million or 3% of worldwide turnover. Details in rights and provenance of AI content.
When in doubt, label. In IAB's January 2026 study, 89% of advertisers using generative AI say they disclose it, but fewer than half always do; and 73% of the young US consumers surveyed say disclosure would increase or not change their purchase intent.
DOOH: what changes on street screens
Production practices we recommend, not statistics.
- Formats: Veo 3.1 and Omni Flash only output 16:9 or 9:16, and Seedance 2.0 spans 21:9 to 9:16. For panoramic screens or LED walls with unusual sizes, generate in the closest standard format and adapt in post.
- Loops: a street asset repeats, so the start and end must join. Veo 3.1's first-and-last-frame control helps, and Midjourney documents loops in its video model.
- No sound: most street screens play no audio; don't pay for it. Silent Veo 3.1 on Google Cloud and silent Kling 3.0 cost less.
- Legibility: slow motion, high contrast and large text composited in post, for an audience that looks for a few seconds from a distance.
- Operator specs: resolution, frame rate, codec, slot length and content policy vary by network; get them before generating.
- Label: the code requires the signal offline too; in a short loop, the simplest option is an icon that stays on screen throughout.
This is the approach we recommend for DOOH. For audience activations, see generative AI brand experiences.
Key takeaways
- Google recommends Gemini Omni Flash as its default video model; Wan 3.0, MiniMax H3 and Seedance share the top of the rankings with it.
- Sora is gone (app on April 26, API on September 24, 2026): archive your masters and keep a second engine tested.
- It works for short spots, social, previs and loops; long continuity, logos and real faces are still human work.
- In our illustrative example, generating a 30 s spot costs $128 to $640; the usable-take rate and human hours drive the budget.
- In the EU, a deepfake is disclosed at first exposure and, in video, at the start and at intervals; in DOOH, a fixed icon is the simplest option.
Sources
- Generate videos with Veo 3.1
- Gemini Omni Flash video generation
- Gemini Developer API pricing
- Deprecations (Videos API and Sora 2)
- Why OpenAI really shut down Sora
- Kling API pricing
- ModelArk pricing (Seedance)
- Wan 3.0: 30-second AI video generation from any input
- Text-to-video leaderboard
- Kalshi airs AI-generated ad during NBA Finals using Google's Veo 3
- Code of Practice on transparency of AI-generated content (final text)
- The AI Ad Gap Widens
Editorial note: this analysis reflects the public information available on the review date. Models, prices and rules change fast; every third-party figure links to its source, and our opinions are labeled as such. Spotted an error? Write to contact@slash-digital.io.
The questions we hear often
What is the best AI video model in 2026?
As of September 26, 2026, Gemini Omni Flash, Wan 3.0, MiniMax H3 and Seedance 2.x lead the preference rankings. For a brand, license, audio, languages and rules on people matter too: test them on your key frames.
What happened to Sora?
OpenAI closed the Sora app and web experience on April 26, 2026 and shut down the API (sora-2 and sora-2-pro) on September 24, 2026, with no replacement.
How much does a second of generated video cost?
Between $0.05 (Veo 3.1 Lite at 720p) and $0.80 (FLUX 3 at 4K) at September 2026 list prices. Multiply by discarded takes: with one usable take in 20, the bill is multiplied by 20.
Does an AI-made ad have to be labeled?
In the EU yes, since August 2, 2026, if it is a deepfake: content that resembles real people, objects, places or events and would falsely appear authentic. Evidently creative works get a lighter regime, not an exemption.
Does generative video work for DOOH screens?
Yes, for short silent loops in 16:9 or 9:16, with text and logos composited in post. Get the operator's specs first and plan the AI icon if the asset runs in the EU.
More analysis to read next
Image models in 2026: which to use for brand work
Nano Banana 2 and Pro, GPT Image 2.5, Qwen-Image: which model to use for brand assets, what each image costs and what the AI Act requires since August 2026.
Image, video and audioGenerative AI in live brand experiences and DOOH: what works
AI photobooths, face swaps and DOOH screens: latency, moderation, biometric consent, AI Act labelling and how to measure a live activation.
Image, video and audioRights and provenance of AI content
Who owns AI output? Lawsuits and settlements, vendor indemnities, likeness and voice, AI Act Article 50, C2PA, SynthID and a practical policy for brand teams.
Where to next
Let's put it in production
Tell us your challenge. We reply within 24 business hours with an honest first read: if we can help, we'll say how; if not, we'll say who can.
I reply personally. No endless forms, no canned replies.