公告
先看证据,再决定买不买

频道每天最多 3 条价格异动与中转状态;具体商品请用机器人设置降价/补货提醒。交流群提问请带预算、模型、工具和使用频率。

查看
社群与联系Telegram 群点击加入Telegram 频道每天最多 3 条有效价格情报联系我们tgAIPricedb交流群979789483
返回资讯列表
product

MiniMax 与 Magnific 在旧金山展示 H3 视频生成能力

MiniMax 表示,其与 Magnific 在旧金山举办了围绕 MiniMax H3 的创作者活动,现场涵盖产品演示、行业讨论和实时体验。

94% VERIFIED

MiniMax 表示,旧金山活动吸引了超过场地容量的观众。MiniMax AI 解决方案架构师 Ethan Wei 现场介绍了 H3,创作者 Enrique Lopez 还在 Magnific 中进行了实时演示。

活动还安排了围绕 AI 视频与创意工作流发展的炉边对谈,Claire Xue 参与讨论。MiniMax 官方资料显示,H3 是一款开放权重的通用多模态生成模型,可结合文本、图像、视频和音频生成最高 2K、最长 15 秒并带原生立体声的视频。

MiniMax 同时称,Magnific 上的 H3 2K 生成服务在 9 月 1 日前可享受五折优惠。该优惠具有明确的时间限制,实际可用性仍应以 Magnific 页面为准。

来源证据

GitHub - MiniMax-AI/MiniMax-H3github.com · supporting

## System Overview MiniMax H3 is a general-purpose, omni-modal generative system. It supports unified understanding of multimodal contexts composed of text, images, video, and audio, and can generate video with native stereo audio at resolutions up to 2K and durations of up to 15 seconds. Thanks to its task-generalization-oriented system design, H3 already possesses broad multimodal context understanding and generation capabilities at the pre-training stage, enabling outstanding performance in following complex multimodal instructions. H3 supports the following input and output specifications:

MiniMaxAI/MiniMax-H3 - vLLM Recipesrecipes.vllm.ai · supporting

vLLMvLLM/Recipes MiniMax # MiniMaxAI/MiniMax-H3 Open-weight general-purpose multimodal generation model — jointly generates 24 FPS video with native stereo audio from text, image, video, and audio references, served via vLLM-Omni 8.7 s of 1248×768 video with synchronized stereo audio in ~87 s on 4×B300 View on HuggingFace dense64B0 ctxvLLM 0.26.0+vLLM-Omninightly path — nightly wheels")omni Guide ## Overview MiniMax H3 is an open-weight, general-purpose multimodal generation model. Rather than being confined to one specialized task — generate, edit, or reference — H3 reads a multimodal context that mixes text, images, video, and audio together, interprets the creative intent as a whole, and produces coherent audio-visual output end to end. [...] The modular root service loads both DiTs by default. Capacity profiles use `--task-type fl2va` or `--task-type ref2va` and expose only that task family. H3 executes one generation request per diffusion batch today. The first request

MiniMax's Postlinkedin.com · supporting

Join Lovis Odin, Creative Engineer at fal, Ethan Wei, AI Solutions Architect ... walkthrough of how creators and developers are building with H3.

MiniMax H3: An Open Model Breaking the Boundaries Between ...minimax.io · supporting

AIH3MultimodalVideo Generation Today, we're launching MiniMax H3, a general-purpose multimodal generation model. H3 understands unified context across text, images, video, and audio, generating video with native stereo sound, up to 15 seconds at 2K resolution. For a quick hands-on experience, please visit . Early testing shows H3 is ready for commercial content creation across a wide range of use cases, excelling at instruction following, accurate text and brand rendering, and V2V motion transfer. With precise, controllable multimodal generation and editing, H3 is built for advertising, branding, e-commerce, product design, UI/UX, gaming, and more. [...] ### Multimodal context understanding Real creative work requires blending complex information across modalities, pulling in images, audio, video, and more as input sources. For example, the prompt for the shot below is: "Reference the Hitchcock camera movement from Video 1, have the character in Image 2 sing, with the vocals matchin

MINIMAX H3 JUST GOT EVEN BETTER!youtube.com · supporting

that it is much faster to first generate the video at a low resolution and then upscale it to a higher resolution at the end. So like for example, if I use the normal text to video workflow with a normal prompt and I use something like 1.2 megapixel resolution and then I click run, which gives us something like this. Yo, Mr. White. Yo, Jesse. What's up? Yeah, there you go. Now right now I'm renting a 4090. So this video took around 172 seconds to generate, which is already fairly fast. And now if I use the fast text to video workflow with the same exact parameters, but this time I first generate the video at 0.2 megapixel and then upscale it to 1.2 megapixel. If now I click run, okay, it gives us something like this at the end. Yo, Mr. White. Yo, Jesse. What's up? So as you can see, first [...] you to extend an already existing video. And I have tried a lot of different workflows, a lot of different nodes, and in my testing, this one is definitely the best. Now, the way it works is ver

We went a little over capacity at @magnific SF for MiniMax ...x.com · supporting

Ethan Wei, AI Solutions Architect at @MiniMax_AI, gave an H3 walkthrough 🛠️ Claire Xue joined us for a fireside on where AI video + creative