MiniMax, the Beijing-based artificial intelligence startup, unveiled its latest video generation model, H3, on July 31, 2026, intensifying an already fierce contest for dominance in AI-generated video. The release positions MiniMax directly against ByteDance, which launched its own updated Seedance 2.5 model the same week in what Bloomberg described as a "dueling" showdown between two of China's most ambitious AI companies. For continuous coverage of the latest AI developments, follow our live reporting.

What H3 brings to the table

According to Reuters, MiniMax's H3 model can generate video clips at 2K resolution — a notable step up from the 1080p output that has been the de facto standard for most consumer-facing video AI tools. The model also produces stereo sound synchronized with the generated footage, a capability that remains rare in the video generation space, where most models output silent clips that require a separate audio pipeline.

Perhaps the most strategically significant decision is MiniMax's commitment to open weights. The South China Morning Post reported that MiniMax is releasing H3 with open-source model weights and at a price point deliberately undercutting competitors — a combination the publication framed as a direct challenge to ByteDance's closed, commercial approach. Open weights allow developers and researchers to download, inspect, fine-tune, and deploy the model on their own infrastructure rather than relying solely on a hosted API.

MiniMax described H3 as an "all-modal" model, according to AIBase, suggesting it is designed to handle multiple types of input and output — text, image, audio, and video — within a single architecture rather than requiring separate specialized models chained together.

Market reaction

Investors responded swiftly. Finance.biggo.com reported that MiniMax's shares surged nearly 13% following the H3 announcement, reflecting market confidence that the company's aggressive pricing and open-weight strategy could capture developer mindshare in a category where ByteDance's Doubao ecosystem has held a commanding lead.

The price war is real. The SCMP noted that MiniMax is offering H3-generated video at a cost substantially below prevailing rates, betting that lower per-generation prices will attract high-volume users — from social media creators to marketing agencies to independent filmmakers — who are highly sensitive to per-clip costs.

ByteDance fires back with Seedance 2.5

ByteDance did not let the moment pass unanswered. On the same timeline, it released Seedance 2.5, the latest iteration of its video generation model. According to Pandaily and CineD, Seedance 2.5 can produce 30-second videos in a single generation — a meaningful improvement over earlier models that typically required stitching together shorter clips to achieve longer sequences.

CineD reported that Seedance 2.5 accepts up to 50 reference inputs, allowing users to guide generation with multiple images, text prompts, and visual references simultaneously. The model also introduces 3D camera blockouts, giving creators director-level control over virtual camera movement — pans, dollies, and orbit paths — within generated scenes. TechNode and Tech in Asia confirmed the model's multi-input and editing capabilities, positioning it as a tool for more sophisticated, narrative-driven video production.

CNET had previously reported that Seedance 2.5 was expected to launch "as soon as this week," and the simultaneous release with MiniMax's H3 underscores how compressed the competitive cycle has become in Chinese AI video.

Why the video AI race matters

The MiniMax-versus-ByteDance rivalry is more than a regional skirmish. AI video generation has become one of the most capital-intensive and technically demanding frontiers in artificial intelligence, requiring enormous compute resources for training and inference. Global competitors — including OpenAI's Sora, Google's Veo, and a growing field of startups — are all racing to produce longer, higher-resolution, more controllable video from text and image prompts.

China's AI video ecosystem has moved with particular speed. MiniMax's decision to open-source H3's weights follows a broader pattern among Chinese AI labs — including DeepSeek, Qwen-maker Alibaba, and others — of releasing powerful open-weight models to build developer ecosystems and establish technical credibility, even at the cost of near-term commercial exclusivity.

The inclusion of stereo audio in H3 is especially notable. Most video generation models today produce visual output only, leaving audio as a separate, often manual step. If MiniMax has genuinely solved synchronized audio generation at 2K resolution, it would represent a meaningful integration milestone — though independent benchmarks and hands-on evaluations will be needed to assess the actual quality and reliability of the output.

The open-weight question

MiniMax's open-weight release also reignites a debate that has split the AI industry. Proponents argue that open weights accelerate research, enable safety scrutiny, and lower barriers for smaller developers. Critics — including some Western frontier labs — contend that releasing powerful generative models openly increases the risk of misuse, from deepfakes to disinformation.

For now, MiniMax is betting that the developer goodwill and ecosystem lock-in from open weights will outweigh any competitive downside. Whether that bet pays off will depend on adoption: if H3 becomes a default tool for video creators across China and beyond, the open-weight strategy could prove as consequential for video AI as Meta's Llama releases were for text.

The immediate competition, though, is clear. Two of China's largest AI companies have fired their latest shots in the video generation wars on the same day — and the pace is only accelerating.

---

Stay Ahead of AI

Get the latest AI news, analysis, and breakthroughs — all in one place.

**Read more AI news →