Maestro v1.5.5 Adds Support for MiniMax H3
Maestro v1.5.5 now supports MiniMax H3, as developers reportedly brought the open-weight model to previously untested gaming GPUs and MacBooks within 48 hours.
Maestro v1.5.5 has added support for MiniMax H3, according to MiniMax and related ecosystem reports. H3 is presented as an open-weight, general-purpose multimodal generation model that accepts text, images, video, and audio, and can produce video up to 15 seconds long at 2K resolution with native stereo sound. These capabilities are primarily vendor claims and have not yet been fully validated by independent testing.
Community developers reportedly ran H3 on hardware MiniMax had not tested, including a gaming GPU and MacBooks operating fully offline. Other projects also added integrations around the release, including ComfyUI, Diffusers, WanGP, and MLX-based tools for Mac systems.
Some users describe local H3 as free or comparable to commercial video models. Those statements are subjective and should not be treated as benchmark results. Actual deployment depends on hardware, memory, software configuration, and the applicable model license.
Source evidence
MiniMaxAI/MiniMax-H3huggingface.co · supporting# MiniMax H3 ## System Overview MiniMax H3 is a general-purpose, omni-modal generative system. It supports unified understanding of multimodal contexts composed of text, images, video, and audio, and can generate video with native stereo audio at resolutions up to 2K and durations of up to 15 seconds. Thanks to its task-generalization-oriented system design, H3 already possesses broad multimodal context understanding and generation capabilities at the pre-training stage, enabling outstanding performance in following complex multimodal instructions. H3 supports the following input and output specifications: [...] To reduce the computational cost of long multimodal sequences, H3 natively supports sparse-attention training and inference. The initial open-source release provides inference with full attention only. Our sparse-attention implementation will be released in a future update. [...] The complete H3 system consists of the following three modules:
MiniMax H3 draws local apps and training tools within 48 hours - RuntimeWireruntimewire.com · supportingYan Junjie, MiniMax's founder, chairman, CEO and CTO, saw outside developers run H3 on untested hardware and build new training, optimization and deployment tools within 48 hours of its open-weights release. On August 4, a day after MiniMax published H3's downloadable checkpoints, MiniMax said Maestro v1.5.5 had added H3 support. MiniMax also said developers had run the model on hardware it had never tested, spanning a gaming GPU and MacBooks operating fully offline. MiniMax on X [...] Maestro was one piece of a larger burst of work documented in MiniMax's ecosystem graphic. MiniMax listed native support in ComfyUI and Diffusers on the day the weights shipped. It said WanGP v12.41 followed with a path designed to run in 5 GB to 6 GB of VRAM, while Phosphene and a MiniMax-H3-MLX engine brought one-click local execution to Macs. DiffSynth-Studio added an NF4 build that MiniMax says lowers the requirement to 7 GB to 8 GB of VRAM. Those memory figures come from MiniMax's own summary rat
MiniMax H3: An Open Model Breaking the Boundaries ...minimax.io · supportingAIH3MultimodalVideo Generation Today, we're launching MiniMax H3, a general-purpose multimodal generation model. H3 understands unified context across text, images, video, and audio, generating video with native stereo sound, up to 15 seconds at 2K resolution. Early testing shows H3 is ready for commercial content creation across a wide range of use cases, excelling at instruction following, accurate text and brand rendering, and V2V motion transfer. With precise, controllable multimodal generation and editing, H3 is built for advertising, branding, e-commerce, product design, UI/UX, gaming, and more. [...] Closed-source models have long dominated video generation, with slower iteration and a less open ecosystem than fields like large language models. To support the open-source community, accelerate compatibility with a broader range of AI hardware, and make it easier for users to build their own customized versions, we plan to open up the model weights in the coming days, subject to
MiniMax-H3 now on huggingfacereddit.com · supportingMy Model is on the second page of Huggingface!Image 73 r/LocalLLM•5mo ago ### My Model is on the second page of Huggingface! 107 upvotes ·44 comments minimax just dropped m3 weights on huggingface. 428b total but only 23b active. anyone tried running it locally yetImage 74 r/ollama•1mo ago ### minimax just dropped m3 weights on huggingface. 428b total but only 23b active. anyone tried running it locally yet Image 75: r/ollama - minimax just dropped m3 weights on huggingface. 428b total but only 23b active. anyone tried running it locally yet 327 upvotes ·51 comments Image 76: Llama Image 77: Llama Public Anyone can view, post, and comment to this community 0 0 [...] Image 27: Claude + GPT + local. Free. Image 28: Your data stays. So does your money. Image 29: Switch models mid-sentence. Free. Image 30: Every model. One app. Free. Image 31: Mage Lab. Local. Free. Your data stays with you. magelab.ai Download is enough for th
MiniMax H3 Unifies AI Video—and Fooocus Has an RCE Warning | AI Signalyoutube.com · supporting[music] AI video models usually make you choose a lane, text to video, image to video, motion reference, editing, or audio. MiniMax says its new H3 model is designed to put those jobs into one system. The company launched H3 on July 31st as a general-purpose multi-model generator. It can take text, images, video, and audio as context, then generate video with native stereo sound for up to 15 seconds at 2K resolution. That could let a creator reference the motion from one clip, the character from an image, and the voice from an audio file in a single instruction. MiniMax also says H3 can handle native multi-shot video, accurate [music] text and brand rendering, motion transfer, and generalized editing. Those are company claims, and the full technical report has not been published yet. The [...] is a verified patched release, do not import metadata from untrusted images. Keep Fooocus off the public internet. Avoid exposing its web interface and run unfamiliar files in an isolated environ
Blaine Brown on X: "Running MiniMax H3 locally for FREE is blowing my mind! It's like having an unrestricted Seedance2.0/Sora2-level model you can run locally for FREE. https://t.co/dmqXrv8z6r" / Xx.com · supporting## Post ## Post user avatar user avatar user avatar user avatar user avatar ## Log in or sign up for X See what’s happening and join the conversation ## Relevant people Avatar ## Trending now