Chinese lab MiniMax’s H3 Max video model can generate a five second, 768p video clip in under three seconds, according to a September 1, 2026 report from QbitAI. Developers have already used the model to run AI-generated livestreams where content is created on the fly.
This article aggregates reporting from 2 news sources. The TL;DR is AI-generated from original reporting. Race to AGI's analysis provides editorial context on implications for AGI development.
MiniMax’s H3 Max pushing video generation below real time is a meaningful inflection point for AI media. Once models can reliably render video faster than it plays, they can power live formats that were previously impossible: continuous AI channels, reactive ad inventory, or interactive experiences where user input changes the scene on the fly. The QbitAI report shows developers already wiring H3 Max into always-on livestreams and TikTok style feeds, not just offline content tools.
For the AGI race, this underscores how quickly Chinese labs are closing gaps in high compute domains like video. Even if H3 Max trails Sora class models in certain cinematic metrics, being ‘good enough’ at real-time generation opens a different set of commercial and research opportunities. It also creates new demand for compute and bandwidth, reinforcing the feedback loop between model capability, infrastructure buildout and energy use.
Strategically, real-time video generation blurs the boundary between model and medium. When the cost of spinning up a new ‘channel’ is near zero, attention becomes the main scarce resource. That will drive more aggressive use of agents to optimize content, targeting and A/B tests in real time. In other words, advances in generative video don’t just enrich the canvas, they incentivize more autonomous decision-making around what people see and when.