Contents

52 / 153

Video Models, by Company

The mid-2026 shift: OpenAI exits consumer video, Google's Omni rebrand, and the Chinese surge

Chapter 51

2 min read

Reviewed v78 · August 2026

The video-model story since version 11 is one of realignment at the top, and the headline is who left. This section records the landscape as of late July 2026.

OpenAI exited consumer video. The Sora consumer app went dark on April 26, 2026, and OpenAI confirmed the Sora API will be discontinued on September 24, 2026. The reported causes are brutal and instructive: operating costs around a million dollars a day against a small fraction of that in revenue, downloads that fell from a 3.3-million peak to roughly a million, a collapsed Disney licensing deal, and unresolved copyright and deepfake problems. Sora continues only as an internal 'world model' research effort. The first major lab to ship a flagship generative-video product became the first to retire one.

Google rebranded rather than iterated. At Google I/O in May, Google did not ship a 'Veo 4.' It announced Gemini Omni, an any-to-any multimodal family with conversational, multi-turn editing and physics-aware generation, and shipped the first model, Gemini Omni Flash, the same day to the Gemini app, Google Flow, and free inside YouTube Shorts. As of July, Gemini Omni Flash sat at number one across the major Artificial Analysis video arenas for text-to-video and image-to-video.

The Chinese labs surged. Kuaishou's Kling 3.0 Turbo, on June 17, added fast previews, multi-shot prompting, and native audio with lip-sync in five languages, and the company raised a round reported to value it around 18 billion dollars, with Alibaba and Tencent among the backers. ByteDance's Seedance 2.5, announced June 23 with API access in mid-July, generates native single-pass 30-second clips with no stitching, accepts up to 50 reference inputs, and outputs 4K. And an Alibaba unit's model, released under the name HappyHorse, climbed to the top of the open video arenas by generating audio and video jointly in a single forward pass.

The reliability caveat, restated. This corner of the field is now thick with search-optimized content asserting releases that never happened: a 'Veo 3.2,' an open-weight 'Wan 3.0,' a mis-dated 'Gen-4.' Treat any video-model claim without an official source or a benchmark entry as unconfirmed. The confirmed picture is enough.

xAI also entered in force. Grok Imagine Video 1.5, launched May 31, 2026, generates short clips with native audio and briefly topped the Artificial Analysis image-to-video arena, a reminder that the video race is no longer just the incumbents and the Chinese labs.

Check your understanding

pass: 5 of 7

Answer at least 5 of 7 correctly to unlock the next chapter.

  1. 1. What did OpenAI do with its Sora consumer video app in 2026?

  2. 2. What were Sora's reported operating costs against its revenue?

  3. 3. Instead of shipping a 'Veo 4,' what did Google announce at Google I/O in May?

  4. 4. What did Kuaishou's Kling 3.0 Turbo add on June 17?

  5. 5. What distinguishes ByteDance's Seedance 2.5?

  6. 6. What made Alibaba's HappyHorse model notable in the open video arenas?

  7. 7. How did xAI enter the video race in 2026?

The weekly briefing

Get the week's moves in your inbox.

A short, sourced digest of what actually moved across generative AI, every week. Free.

Free. One email a week, no spam, unsubscribe anytime. Prefer a reader? RSS.