Объявление
数据公告

QQ群和tg群已经启用,欢迎加入。公开信息来源均审核后发布;请结合来源、库存和更新时间判断。

Сообщество и контактыTelegram 群点击加入Telegram 频道点击订阅联系我们tgAIPricedb交流群979789483
К списку новостей
Продукты

MiniMax представила мультимодальную видеомодель H3 с открытыми весами

MiniMax запустила H3 — видеомодель с открытыми весами, которая объединяет обработку текста, изображений, видео и аудио. По данным компании, она генерирует ролики длительностью до 15 секунд в разрешении 2K со встроенным стереозвуком.

92% VERIFIED

MiniMax позиционирует H3 как универсальную мультимодальную модель для коммерческого создания контента. Она должна понимать текст, изображения, видео и аудио в едином контексте, а затем создавать аудиовизуальный результат с нативным стереозвуком. Среди целевых направлений компания называет рекламу, брендинг, электронную коммерцию, дизайн продуктов, UI/UX и игры, отдельно отмечая следование инструкциям, отображение текста и брендов, а также перенос движения между видео.

H3 опубликована с открытыми весами и доступна через региональные API и приложения MiniMax, а также через некоторые сторонние платформы. В описаниях продукта упоминаются генерация видео из текста, работа с первым и последним кадром, создание видео по референсам и точное редактирование видео.

Первые отзывы пользователей в целом отмечают согласованность изображения, качество звука и высокую скорость. При этом локальный запуск может требовать производительного оборудования, а на некоторых системах длительный рендеринг оказывается нестабильным. Пользователи также сообщают, что модель иногда добавляет музыку или звуковые эффекты без явного запроса, поэтому для рабочих сценариев важны точные инструкции и контроль параметров.

Источники

MiniMax H3 - Open-Weights General-Purpose Multimodal ...fal.ai · supporting

## Common questions about MiniMax H3 MiniMax H3 is an open-weights, general-purpose multimodal video model. Instead of a separate model for each task, MiniMax H3 reads text, images, video, and audio in one unified context and generates coherent audiovisual results from any mix of them. It supports text-to-video, first-and-last-frame, reference-to-video, and precise video editing. MiniMax H3 is released with open weights, so it is an open foundation you can explore, customize, and build on rather than a closed endpoint. fal is a Day 0 ecosystem partner, which means you can call the hosted MiniMax H3 API on fal.ai from launch without provisioning GPUs, and still have the option to work with the weights directly for your own research and fine-tuning. [...] ### Green-Screen Fairytale Composite "Remove the green screen background of Video 1 and turn it into a fairy tale-like background similar to Video 2. The background elements need to completely match the actions of the characters in V

What Is MiniMax H3 (Hailuo 3.0)? The Open-Weight ...huggingface.co · supporting

MiniMax H3 (Hailuo 3.0) launched July 31, 2026 as a general-purpose omni-modal generation model: one transformer that understands text, images, video, and audio together and returns video with native stereo sound at up to 2K and 15 seconds. Here is the architecture, the API, the pricing, the open-weights reality — and whether you should build on it today. August 1, 2026 · ~10 min read minimax-h3-hero minimax-h3-hero [...] MiniMax H3 is the official model name; Hailuo 3.0 — also written Hailuo 03 — is the widely used alias, after the Hailuo AI app it ships in. They are one model, the direct successor to Hailuo 2.3. MiniMax announced it on July 31, 2026 as "a general-purpose omni-modal generation model" that "jointly understand[s] multimodal contexts spanning text, images, video, and audio." Output is video with native stereo audio at up to 2K and 15 seconds. Morphic's model page confirms the naming — and warns it is not related to Kling O3, a Kuaishou model the "Hailuo 03" spelling i

MiniMax H3: The Open-Weights Video Model That Just Undercut Everyone by 3x | SaaSCitysaascity.io · supporting

MiniMax released H3 on July 31, 2026 — a general-purpose omni-modal generation model that takes text, images, video, and audio in a single context and returns 2K video with native stereo sound. The Hong Kong market noticed immediately: shares of MiniMax (0100.HK) climbed roughly 13% by midday, touching HK$231.6 intraday, with nearly HK$1 billion in turnover on the session. But the stock move isn't the point. The point is that video generation has been the last major modality where closed models had a comfortable moat, and MiniMax just announced it's walking away from that moat on purpose. ## What MiniMax H3 Actually Does [...] The commercial targeting is explicit. MiniMax is aiming this at advertising, e-commerce, branding, product design, UI/UX, and gaming — not at people making surreal TikToks. Film title sequences, game menu animations, motion posters, product reveals. ## The Omni-Reference Trick Is the Real Unlock Most video pipelines today are Frankenstein jobs. You generate a

MiniMax H3: An Open Model Breaking the Boundaries Between Tasks and Modalities | MiniMax | 10 commentslinkedin.com · supporting

Agree & Join LinkedIn By clicking Continue to join or sign in, you agree to LinkedIn’s User Agreement, Privacy Policy, and Cookie Policy. # MiniMax’s Post View organization page for MiniMax 32,304 followers MiniMax H3: Omni-Reference, Commercial-Grade Generation, Unbeatable Cost Efficiency, Open Weights MiniMax H3: An Open Model Breaking the Boundaries Between Tasks and Modalities Bjørn Furuknap, graphic Very impressive model, although it tends to be very eager to add audio and music, even when told not to. It's also fast. Very fast. If you have the right hardware. My laptop 3070 render crashed after 13 hours on a 5-second clip. The 5090 does the same in 120 seconds. I made a guide to set it up in about 10 minutes on ComfyUI using rented 5090s at Runpod: Ainomíra Ilario, graphic [...] Marcelo Javier Dayer de Zehnder, graphic This model is impressive. The level of consistency, sound, and image quality really shows the great work in the lab. Kudos to you and the team. This is a

MiniMax H3: An Open Model Breaking the Boundaries ...minimax.io · supporting

AIH3MultimodalVideo Generation Today, we're launching MiniMax H3, a general-purpose multimodal generation model. H3 understands unified context across text, images, video, and audio, generating video with native stereo sound, up to 15 seconds at 2K resolution. Early testing shows H3 is ready for commercial content creation across a wide range of use cases, excelling at instruction following, accurate text and brand rendering, and V2V motion transfer. With precise, controllable multimodal generation and editing, H3 is built for advertising, branding, e-commerce, product design, UI/UX, gaming, and more. [...] MiniMax M3MiniMax M2.7MiniMax M2.5 VIDEO MiniMax H3 SPEECH & MUSIC MiniMax Speech 2.8MiniMax Music 3.0 Product MiniMax CodeMiniMax HubAudioTalkie API Token Plan Research Company Intelligence with everyone AboutNewsInvestor Relations Contact Us Models LLM MiniMax M3NEWMiniMax M2.7MiniMax M2.5 VIDEO MiniMax H3NEW SPEECH & MUSIC MiniMax Speech 2.8NEWMiniMax Music

MiniMax (official) on X: "MiniMax H3: Omni-Reference, Commercial-Grade Generation, Unbeatable Cost Efficiency, Open Weights https://t.co/DLB1xsFfSC" / Xx.com · supporting

Log inSign up ## Post user avatar MiniMax (official) @MiniMax\_AI MiniMax H3: Omni-Reference, Commercial-Grade Generation, Unbeatable Cost Efficiency, Open Weights Article cover image # MiniMax H3: An Open Model Breaking the Boundaries Between Tasks and Modalities 1:53 AM · Jul 31, 2026431.6KViews user avatar MiniMax (official) @MiniMax\_AI 8h Online API Use MiniMax-H3 directly via API. Global: platform.minimax.io | CN: platform.minimaxi.com Online App Use MiniMax-H3 directly via App. WebApp Global: hailuoai.video | CN: hailuoai.com Desktop Global: hub.minimax.io | CN: Models - MiniMax API DocsFrom platform.minimax.io 2K user avatar RyanLee MiniMax (official) @RyanLeeMiniMax Jul 31 user avatar RyanLee MiniMax (official) [...] Jul 31 user avatar RyanLee MiniMax (official) @RyanLeeMiniMax Jul 31 As I said, MiniMax-H3 is Open! #1 Video Editing (With Audio) #2 Text to Video (With Audio) #2 Image to Video (N