Notice
数据公告

QQ群和tg群已经启用,欢迎加入。公开信息来源均审核后发布;请结合来源、库存和更新时间判断。

Community & contactTelegram 群点击加入Telegram 频道点击订阅联系我们tgAIPricedb交流群979789483
Back to news
Products

MiniMax Launches H3 Open-Weight Multimodal Video Model

MiniMax has introduced H3, an open-weight video model designed to process text, images, video, and audio in a unified context. The company says it can generate up to 2K, 15-second video with native stereo sound.

92% VERIFIED

MiniMax describes H3 as a general-purpose multimodal generation model for commercial content workflows. It is intended to understand text, images, video, and audio together, then produce audiovisual output with native stereo sound. The company highlights instruction following, text and brand rendering, video-to-video motion transfer, and use cases spanning advertising, e-commerce, product design, UI/UX, and gaming.

The model is distributed with open weights and is also available through MiniMax’s regional APIs, applications, and selected ecosystem platforms. Public descriptions list text-to-video, first-and-last-frame generation, reference-to-video, and video-editing workflows among its supported capabilities.

Early community feedback is broadly positive about visual consistency, sound quality, and speed, but suggests that local inference can demand powerful hardware and may be unreliable on less capable systems. Users have also reported that H3 can add music or sound effects even when they are not wanted, making prompt and workflow controls important for production use.

Source evidence

MiniMax H3 - Open-Weights General-Purpose Multimodal ...fal.ai · supporting

## Common questions about MiniMax H3 MiniMax H3 is an open-weights, general-purpose multimodal video model. Instead of a separate model for each task, MiniMax H3 reads text, images, video, and audio in one unified context and generates coherent audiovisual results from any mix of them. It supports text-to-video, first-and-last-frame, reference-to-video, and precise video editing. MiniMax H3 is released with open weights, so it is an open foundation you can explore, customize, and build on rather than a closed endpoint. fal is a Day 0 ecosystem partner, which means you can call the hosted MiniMax H3 API on fal.ai from launch without provisioning GPUs, and still have the option to work with the weights directly for your own research and fine-tuning. [...] ### Green-Screen Fairytale Composite "Remove the green screen background of Video 1 and turn it into a fairy tale-like background similar to Video 2. The background elements need to completely match the actions of the characters in V

What Is MiniMax H3 (Hailuo 3.0)? The Open-Weight ...huggingface.co · supporting

MiniMax H3 (Hailuo 3.0) launched July 31, 2026 as a general-purpose omni-modal generation model: one transformer that understands text, images, video, and audio together and returns video with native stereo sound at up to 2K and 15 seconds. Here is the architecture, the API, the pricing, the open-weights reality — and whether you should build on it today. August 1, 2026 · ~10 min read minimax-h3-hero minimax-h3-hero [...] MiniMax H3 is the official model name; Hailuo 3.0 — also written Hailuo 03 — is the widely used alias, after the Hailuo AI app it ships in. They are one model, the direct successor to Hailuo 2.3. MiniMax announced it on July 31, 2026 as "a general-purpose omni-modal generation model" that "jointly understand[s] multimodal contexts spanning text, images, video, and audio." Output is video with native stereo audio at up to 2K and 15 seconds. Morphic's model page confirms the naming — and warns it is not related to Kling O3, a Kuaishou model the "Hailuo 03" spelling i

MiniMax H3: The Open-Weights Video Model That Just Undercut Everyone by 3x | SaaSCitysaascity.io · supporting

MiniMax released H3 on July 31, 2026 — a general-purpose omni-modal generation model that takes text, images, video, and audio in a single context and returns 2K video with native stereo sound. The Hong Kong market noticed immediately: shares of MiniMax (0100.HK) climbed roughly 13% by midday, touching HK$231.6 intraday, with nearly HK$1 billion in turnover on the session. But the stock move isn't the point. The point is that video generation has been the last major modality where closed models had a comfortable moat, and MiniMax just announced it's walking away from that moat on purpose. ## What MiniMax H3 Actually Does [...] The commercial targeting is explicit. MiniMax is aiming this at advertising, e-commerce, branding, product design, UI/UX, and gaming — not at people making surreal TikToks. Film title sequences, game menu animations, motion posters, product reveals. ## The Omni-Reference Trick Is the Real Unlock Most video pipelines today are Frankenstein jobs. You generate a

MiniMax H3: An Open Model Breaking the Boundaries Between Tasks and Modalities | MiniMax | 10 commentslinkedin.com · supporting

Agree & Join LinkedIn By clicking Continue to join or sign in, you agree to LinkedIn’s User Agreement, Privacy Policy, and Cookie Policy. # MiniMax’s Post View organization page for MiniMax 32,304 followers MiniMax H3: Omni-Reference, Commercial-Grade Generation, Unbeatable Cost Efficiency, Open Weights MiniMax H3: An Open Model Breaking the Boundaries Between Tasks and Modalities Bjørn Furuknap, graphic Very impressive model, although it tends to be very eager to add audio and music, even when told not to. It's also fast. Very fast. If you have the right hardware. My laptop 3070 render crashed after 13 hours on a 5-second clip. The 5090 does the same in 120 seconds. I made a guide to set it up in about 10 minutes on ComfyUI using rented 5090s at Runpod: Ainomíra Ilario, graphic [...] Marcelo Javier Dayer de Zehnder, graphic This model is impressive. The level of consistency, sound, and image quality really shows the great work in the lab. Kudos to you and the team. This is a

MiniMax H3: An Open Model Breaking the Boundaries ...minimax.io · supporting

AIH3MultimodalVideo Generation Today, we're launching MiniMax H3, a general-purpose multimodal generation model. H3 understands unified context across text, images, video, and audio, generating video with native stereo sound, up to 15 seconds at 2K resolution. Early testing shows H3 is ready for commercial content creation across a wide range of use cases, excelling at instruction following, accurate text and brand rendering, and V2V motion transfer. With precise, controllable multimodal generation and editing, H3 is built for advertising, branding, e-commerce, product design, UI/UX, gaming, and more. [...] MiniMax M3MiniMax M2.7MiniMax M2.5 VIDEO MiniMax H3 SPEECH & MUSIC MiniMax Speech 2.8MiniMax Music 3.0 Product MiniMax CodeMiniMax HubAudioTalkie API Token Plan Research Company Intelligence with everyone AboutNewsInvestor Relations Contact Us Models LLM MiniMax M3NEWMiniMax M2.7MiniMax M2.5 VIDEO MiniMax H3NEW SPEECH & MUSIC MiniMax Speech 2.8NEWMiniMax Music

MiniMax (official) on X: "MiniMax H3: Omni-Reference, Commercial-Grade Generation, Unbeatable Cost Efficiency, Open Weights https://t.co/DLB1xsFfSC" / Xx.com · supporting

Log inSign up ## Post user avatar MiniMax (official) @MiniMax\_AI MiniMax H3: Omni-Reference, Commercial-Grade Generation, Unbeatable Cost Efficiency, Open Weights Article cover image # MiniMax H3: An Open Model Breaking the Boundaries Between Tasks and Modalities 1:53 AM · Jul 31, 2026431.6KViews user avatar MiniMax (official) @MiniMax\_AI 8h Online API Use MiniMax-H3 directly via API. Global: platform.minimax.io | CN: platform.minimaxi.com Online App Use MiniMax-H3 directly via App. WebApp Global: hailuoai.video | CN: hailuoai.com Desktop Global: hub.minimax.io | CN: Models - MiniMax API DocsFrom platform.minimax.io 2K user avatar RyanLee MiniMax (official) @RyanLeeMiniMax Jul 31 user avatar RyanLee MiniMax (official) [...] Jul 31 user avatar RyanLee MiniMax (official) @RyanLeeMiniMax Jul 31 As I said, MiniMax-H3 is Open! #1 Video Editing (With Audio) #2 Text to Video (With Audio) #2 Image to Video (N