Notice
数据公告

QQ群和tg群已经启用,欢迎加入。公开信息来源均审核后发布;请结合来源、库存和更新时间判断。

Community & contactTelegram 群点击加入Telegram 频道点击订阅联系我们tgAIPricedb交流群979789483
Back to news
Products

MiniMax-H3 Demonstrates Local Video Generation on One RTX 5090, With Important Caveats

A public demonstration shows MiniMax-H3 generating video locally on a single RTX 5090. However, hardware requirements vary widely by precision, resolution, and workflow, and the complete 2K pipeline is not fully local.

82% VERIFIED

MiniMax has highlighted a demonstration in which H3 generates video locally on one RTX 5090. The model is described as supporting text, image, video, and audio references, including short clips with synchronized stereo audio. Its weights are publicly available, and community integrations are being developed for tools such as ComfyUI.

The single-GPU demonstration should not be interpreted as proof that every H3 configuration runs comfortably or entirely offline on one card. Technical reports describe an original task partition containing roughly 134 GiB of weights, while higher-quality or lossless setups may require several data-center GPUs. Consumer-GPU workflows generally depend on quantization, host-memory or disk offloading, and reduced settings that can affect speed, precision, or resolution.

The full 2K workflow also reportedly relies on API-only components, including Context-IR and Regenerate-2K. The strongest supported conclusion is therefore that some H3 video-generation workflows are now feasible on consumer hardware, not that the complete quality stack is available locally on a single RTX 5090.

Source evidence

MiniMax H3 Review: Benchmarks, Specs and Hardwarekingy.ai · supporting

### Can MiniMax H3 run completely locally? H3-Base can run locally and generates 768p-class video with stereo audio. The official full-2K workflow cannot run completely locally because Context-IR and Regenerate-2K are API-only at launch. ### How much VRAM does MiniMax H3 need? There is no single number. One original task partition contains about 134 GiB of weights. Published lossless recipes include four 80 GB H100s, four H200s and two 32 GB RTX 5090s with roughly 384 GiB host RAM. Diffusers documents INT8 streaming on 24–32 GB cards with about 75 GB host RAM, and reduced-canvas operation on 12–16 GB cards with more offload. Those smaller routes trade speed, precision or both. ### Will MiniMax H3 run on a Mac? [...] Fast interconnect matters. Ulysses sequence parallelism divides the long token sequence but does not automatically shard every weight. Tensor parallelism and FSDP reduce per-card weight residency at the cost of collectives. High-bandwidth single-node NVLink/NVSwitch or

Minimax H3 is quite resource intensive like no model before : r/StableDiffusionreddit.com · supporting

Able-Instruction1009 •11h ago 3060 12GB / 32GB RAM / Ubuntu. Not a CUDA OOM: the whole machine locked, SSH included, reset button both times. Culprit was in the startup log — Enabled pinned memory 30419.0. ComfyUI page-locking ~30GB on a 32GB system. The kernel can’t reclaim pinned memory, so you get swap livelock instead of an OOM kill. Add --disable-pinned-memory (I also run --cache-none --fast-disk --disable-smart-memory --lowvram). Now: 480p, 5s @ 24fps, native audio, one ref image. Peak 7.4GB RAM used and 6.2GB VRAM — nowhere near the ceiling I thought I was hitting. More replies Image 13: u/Full_Astronomer_5438 avatar Full_Astronomer_5438 •16h ago will look into that, thanks! [...] Image 8: u/1or4s avatar 1or4s •16h ago Generating a video now on a 3060 12gb with 48gb of ram. Gpu is at a consistent 75°. 42 gb of ram being used. No activity on the hard drive. I undervolt my card and you should too if you are experiencing high gpu temps. Reply Share Image 9: u/Able

MiniMax-H3 Open Video AI Model Tops Benchmarks on ...x.com · supporting

00:00 253K user avatar MiniMax (official) @MiniMax\_AI 17h This is how much you can do with a single RTX 5090 NOW with MiniMax H3. WE HAVE CROSSED A LINE.✊ user avatar Ryan Lightbourn @ryanlightbourn Aug 3 Generated locally with H3 on an RTX 5090 in a few minutes, from nothing but a text prompt. We have crossed a line. 00:00 88K user avatar Arena.ai @arena 16h Big news: MiniMax-H3 by @MiniMax\_AI is now the #1 open model in Video Arena: across both Text-to-Video and Image-to-Video. This is +280pts over the next best open model, hunyuan-video-1.5, and a huge improvement from Hailuo-2.3 at #27 (1199 pts) and Hailuo-02-pro at #28 (1197 00:41 user avatar MiniMax (official) @MiniMax\_AI Aug 3 MiniMax-H3 Is Now Publicly Available huggingface.co/MiniMaxAI/Mini… [...] Log inSign up ## Trending # MiniMax-H3 Open Video AI Model Tops Benchmarks on Hugging Face Last updated Aug 4, 2026 MiniMax released its 33-billion-parameter MiniMax-H3 model as open weights on Huggi

MiniMax released a 33B video model that generates ...x.com · supporting

Holy, those insanae releases: MiniMax released a 33B video model that generates synchronized stereo audio and runs on one RTX 5090. H3 combines text, images, video and audio references for generation and editing, with clips up to 15 seconds. ComfyUI’s optimized stack is roughly 40GB, using dynamic RAM/SSD offloading to fit consumer GPUs. Early 5090 (!) tests produced five seconds at native 768p-class resolution in around 5.5 minutes. MiniMax calls it open source, but several core pieces remain server-side: context orchestration, 2K regeneration and sparse attention. Open weights. Closed quality stack. Restricted geography. However, H3 is a major step for local video, but also a reminder that “downloadable” does not mean open source in general. Btw: It also cannot legally be used under

MiniMax (official) on X: "This is how much you can do with a single RTX 5090 NOW with MiniMax H3. WE HAVE CROSSED A LINE.✊" / Xx.com · supporting

Log inSign up ## Post user avatar MiniMax (official) @MiniMax\_AI This is how much you can do with a single RTX 5090 NOW with MiniMax H3. WE HAVE CROSSED A LINE.✊ user avatar Ryan Lightbourn @ryanlightbourn Aug 3 Generated locally with H3 on an RTX 5090 in a few minutes, from nothing but a text prompt. We have crossed a line. 00:00 9:16 PM · Aug 3, 2026115.8KViews user avatar 🤦🏻‍♂️TheDUMBESTguyInAi🤦🏻‍♂️ @LeanKinPrazli 20h This is on my 3090 00:00 1.2K user avatar Ryan Lightbourn @ryanlightbourn 21h ty frens, models are keeping me warm a little too warm it's 96 degrees in my office GIF 1.9K user avatar Rafa Schwinger 🇻🇦 @Rafa\_Schwinger 20h [...] Rafa Schwinger 🇻🇦 @Rafa\_Schwinger 20h the skin is plasticky bro, can you please check if the fp16 VAE is overly quantized compared to the OG fp32 VAE? 1.6K

Ryan Lightbourn (@ryanlightbourn) / Xx.com · supporting

Log inSign up Ryan Lightbourn user avatar Ryan Lightbourn Rock Sound Joined August 2009 340 Following11.9K Followers RepliesRepliesMediaMedia Pinned user avatar Ryan Lightbourn @ryanlightbourn Jul 9 I grew up watching Hook🌈🌞 00:00 17K user avatar Ryan Lightbourn @ryanlightbourn 10h Generated locally with H3 on an RTX 5090 in a few minutes, from nothing but a text prompt. We have crossed a line. 00:00 39K user avatar Ryan Lightbourn @ryanlightbourn Jul 30 Who's ready to embark on a quest? H3 is coming. @Hailuo\_AI #MiniMaxH3 00:00 6.3K user avatar Ryan Lightbourn @ryanlightbourn Jul 24 In August I'll be back in the kitchen, cooking up my 3rd short film: user avatar MoonPay 🟣 @moonpay [...] user avatar MoonPay 🟣 @moonpay Jul 24 the Lumara Film Festival is now open for submissions! 🎬 create an original movie with AI 🏆 compete for awards in 7 categories 💰 $100,000 pr