MiniMax H3 可在 RTX 4090 上本地生成视频,社区分享 ComfyUI 工作流
社区用户展示了如何使用 MiniMax H3 和 ComfyUI 在配备 24GB 显存的 RTX 4090 上本地制作视频,并分享提示词编写、节奏同步和图生视频等实践经验。
一名用户分享了使用 MiniMax H3 与 ComfyUI 本地制作 AI 视频的体验,涵盖文本生视频、根据音乐鼓点切换画面,以及图生视频三类工作流。其方法是先整理创意,再借助官方 H3 提示词指南将想法改写为结构化提示词,并通过多轮调整匹配预期画面。
据该用户介绍,RTX 4090 24GB 在 20 步设置下生成 15 秒视频约需 200 秒以上,提升到 768P 后可能超过 1,000 秒;使用官方 API 放大到 2K 约需 6 至 8 分钟。实际速度会受到分辨率、模型量化、步骤数和首次加载时间等因素影响。
相关社区资料显示,H3 已被用于 ComfyUI 本地工作流,官方提示词文档和模板可帮助初学者快速开始。该用户还计划免费发布本次使用的工作流、脚本和素材,但具体仓库内容及性能数据仍应以原作者后续发布为准。
来源证据
No GPU? No Problem. A Complete MiniMax H3 ...medium.com · supporting`nvfp4_awq` `pruned_int8_convrot` `fp8_scaled` First run: start small. Template library → Video → MiniMax H3 → Text to Video (the three `api_`-prefixed templates call the cloud API and need a key — ignore them for now). Check that the UNETLoader points at the FL2VA weight, the encoder is the NVFP4 build, and both VAEs are loaded. A sensible 12 GB starting point: 16:9, 864×480 (~0.4 MP), 5 seconds, Turbo on, Turbo steps 8 — this validates the environment rather than chasing quality. Keep the prompt to one scene, one main action, one shot, and sound. The first generation is always slow (about 34 GB of weights load first; reference: 4090 48G first run ≈ 644 s vs ≈ 425 s warm). `api_` [...] Prompt format. H3 uses an official structured prompt format (full Prompting Guidance in the official README). A practical trick: hand your Chinese/plain brief plus the official prompt guide to any LLM and let it rewrite into the structured format, then save as a template — an effective community stand
[No GPU Required] How I Started Making AI Videos with ...note.com · supportingHow I Started Making AI Videos with Audio Using 'MiniMax H3' you need a VRAM 24GB class GPU. That's the world of the RTX 4090 or 5090. This
MiniMax H3 for ComfyUI achieves impressive results on ...facebook.com · supportingMiniMax H3 for ComfyUI is here! I am getting insane results out of it on just a RTX 4090 laptop. Leaps and bounds if we can get a replacement
Run MiniMax H3 Locally: VRAM Guide From 6GB Cards to ...mindstudio.ai · supportingA rough tiering, based on what’s been demonstrated: 6GB VRAM (RTX 2060 class): workable but degraded, expect mushy detail and rougher audio, generation times around 10-15 minutes for short clips. 12-16GB VRAM (RTX 4060 class): noticeably cleaner, faster generation, still below full quality. 24GB VRAM (RTX 4090 class): enough headroom to also do LoRA training on top of generation. 32GB+ VRAM (RTX 5090, RTX 6000): the intended target, supports longer 15-second clips at higher resolution with less compromise. The pattern lines up with most modern diffusion-based video models: more VRAM buys you resolution, generation length, and speed, but low-VRAM setups aren’t dead ends anymore thanks to aggressive quantization and optimization work from the open-source community. [...] MiniMax H3 is fully open source, meaning anyone can download the weights and run generation locally instead of relying on a paid API. VRAM needs scale from roughly 6GB to 24GB+, with lower-VRAM setups producing so
Running MiniMax H3 on One RTX 4090 - Pedro Alonsopedroalonso.net · supportingThe storybook film took an afternoon once the pipeline worked: six illustrations, three techniques, about 25 minutes of GPU time for the shots plus one reference-mode generation. A year ago this was a render-farm conversation. ## Workflows for download Everything in this post runs on core ComfyUI nodes — v0.30 or later, no custom node packs to install. The weights are on Hugging Face at Comfy-Org/MiniMax-H3, and MiniMax’s structured prompt guides ship in the model repo’s `docs/` folder (`VIDEO_PROMPT_WRITING_GUIDE_base_en.md` and the `_ref_` variant for reference mode). [...] Meanwhile, rendering the same freeform prompt at 5× the pixels changed nothing: 608×352 scored 29.2, 864×480 scored 30.3, 1344×768 scored 26.0. If your generations look soft, the fix is not more resolution. To be fair to the model: the atmospheric prompt asked for shallow depth of field, film grain, sea mist, and volumetric fog — and H3 rendered exactly that. It was doing its job. But the gap between hand-tuned
MiniMax (official) on Xx.com · supportingVery useful resources on making videos with the Minimax H3 model on an RTX 4090 (24GB) 😊 The showcases look amazing, and detailed prompts plus