Notice
先看证据,再决定买不买

频道每天最多 3 条价格异动与中转状态;具体商品请用机器人设置降价/补货提醒。交流群提问请带预算、模型、工具和使用频率。

View
Community & contactTelegram 群点击加入Telegram 频道每天最多 3 条有效价格情报联系我们tgAIPricedb交流群979789483
Back to news
Products

MiniMax Releases Music 3 Open-Weight Model for Songs Up to Five Minutes

MiniMax Music 3 is now available as an open-weight text-to-music model that generates complete songs from lyrics and music descriptions. Its weights are hosted on Hugging Face, with support for several local inference tools.

94% VERIFIED

MiniMax Music 3 is an open-weight text-to-music model designed to turn lyrics and a detailed musical description into a complete song. It can generate tracks of up to about five minutes, including vocals and evolving arrangements, and outputs 32 kHz, 16-bit stereo WAV audio.

According to MiniMax, the system uses a hybrid architecture with an 8B Global LLM for long-range semantic and structural planning and a 0.6B Local LLM for frame-level acoustic detail. The company also describes continuous hidden-state conditioning together with Flow Matching and Flow-VAE components intended to improve vocal rendering and long-form stability.

The model weights are available on Hugging Face, with documented deployment paths involving SGLang, diffusers, and ComfyUI. A hosted cloud API is described as coming soon, so the currently documented route is local inference rather than an already available public API.

Source evidence

MiniMax open-sources Music 3, a Qwen3-based song generator | AI Weeklyaiweekly.co · supporting

## Shared on Bluesky by 2 AI experts Sung Kim @sungkim.bsky.social: 🎵MiniMax-Music3 (open-weight) Next-Generation Open-Weights Production-Ready & Versatile Music Model Model: huggingface.co/MiniMaxAI/Mi... … → M MiniMax (official): huggingface: github: 魔搭: → Originally reported by github.com Read the original article → Original headline: GitHub - MiniMax-AI/MiniMax-Music3 Free AI alerts in your inbox Breaking AI news 3x/week. 50,000+ subscribers. We use essential cookies to keep the site working (login, form security). With your permission, we also use analytics cookies to understand how you use the site. Privacy policy [...] MiniMax has open-sourced MiniMax Music 3, a text-to-song model that generates complete tracks up to five minutes long from lyrics plus a music description. The output is 32 kHz, 16-bit stereo WAV, and lyrics can carry section tags like [Intro], [Verse], [Pre-Chorus], [Chorus], [Bridge], [Instrumental], [Solo] and [Outro] to shape structure.

MiniMaxAI/MiniMax-Music3huggingface.co · supporting

# MiniMax Music 3 MiniMax Music 3 is a high-performance music generation model for creating complete songs up to five minutes long. Conditioned on lyrics and a detailed music description, it generates structurally coherent songs with expressive vocals, evolving arrangements, and stable long-form audio quality. MiniMax Music 3 combines an 8B Global LLM for long-range musical structure, a 0.6B Local LLM for frame-level acoustic detail, and a continuous hidden-state synthesis system based on Flow Matching and Flow-VAE. The model produces 32 kHz, 16-bit stereo WAV audio. ## Demo Explore music generation examples on the MiniMax Music 3 Demo. ## Complete Songs with Long-Range Coherence [...] ### Download the Model `hf download MiniMaxAI/MiniMax-Music3 --local-dir /path/to/minimax_ttm` We recommend the following inference frameworks to serve the model: SGLang - see cookbook diffusers - see diffusers docs ComfyUI see comfyUI tutorials ### Serve with SGLang-Omni `sgl-omni serve --mo

What Is MiniMax Music 3? Open-Weights Music Modelkie.ai · supporting

MiniMax Music 3 is an open-weights AI music generation model from MiniMax that turns lyrics and a text music description into a complete, produced song of up to five minutes in a single generation. It outputs 32 kHz, 16-bit stereo audio and went live on August 13, 2026, with weights on Hugging Face and native support in ComfyUI. MiniMax's official announcement called it a "Next-Generation Open-Weights Production-Ready & Versatile Music Model," with a cloud API listed as coming soon. ## Key Takeaways [...] It targets the parts of music generation that short prompts struggle with: understanding a creator's expressive intent, holding that intent across an entire song, rendering instruments with physical realism, and producing vocals that sound performed rather than synthesized. The status is released, not leaked. MiniMax's official blog dated the launch August 13, 2026, and the weights are live on the `MiniMaxAI/MiniMax-Music3` Hugging Face repository, with a ComfyUI-optimized distributi

MiniMax Music 3: The Open-Weight AI Music Model ...mindstudio.ai · supporting

MiniMax Music 3 is an open-weight AI music generator with downloadable weights on Hugging Face, built around a Qwen3-8B language model paired with a diffusion-based audio pipeline. It runs locally with a claimed minimum of 8GB of VRAM using layer streaming, though 20 to 24GB is recommended for smooth full-precision inference. The model already has day-one support in ComfyUI, and community fine-tunes started appearing on Hugging Face almost immediately after release. Its license permits commercial use with attribution to MiniMax Music 3, and only requires a separate agreement with MiniMax once a project earns more than $20 million. [...] ## What is MiniMax Music 3? MiniMax Music 3 is an open-weight text-to-music model released by MiniMax, the same company behind the open-weight video generator Hailuo (also referred to as H3 in some coverage) and the Seance 2.5 model. Music 3 takes a text prompt, and in some workflows a style or lyric reference, and generates a full song with vocals

MiniMax Music 3.0: Next-Generation Open-Weights, Production-Ready & Versatile Music Model - MiniMax Research | MiniMaxminimax.io · supporting

At the model level, Music 3.0 uses a global–local collaborative Hybrid-LM. The 8B Global LLM predicts core semantic and structural tokens frame by frame and maintains full-song context. The 0.6B Local LLM predicts acoustic tokens along the depth axis within each frame, supplying local sonic detail. This division of responsibilities maintains temporal stability across songs of up to five minutes while preserving rich variation within individual sections. [...] Music 3.0 introduces a new audio-rendering system designed to produce more natural, studio-quality vocal performances. The Structured Caption describes vocal timbre, delivery, techniques such as breathiness and falsetto, harmony arrangement, and effects such as delay and Auto-Tune in fine detail. Continuous hidden states fused from the global and local language models carry this performance information into the flow-matching and Flow-VAE generation process. Together, these mechanisms reduce the high-frequency digital artifacts com

🎵MiniMax-Music3 (open-weight) Next-Generation Open-Weights Production-Ready & Versatile Music Model Model: https://huggingface.co/MiniMaxAI/MiniMax-Music3 Repo: https://github.com/MiniMax-AI/MiniMax-Music3threads.com · supporting

Related threads digitalmatters.me's profile picture digitalmatters.me AI\_Music\_Generation MiniMax Music 3.0 Puts a Song Model on Your Own Hardware. Read the License First. MiniMax Music 3 is an open-weight music generation model from the Chinese lab MiniMax, published on August 13, 2026.... digitalmatters.me/artif… AI\_Music\_Generation #Generative\_AI #MiniMax #Open\_Weights MiniMax Music 3.0 Puts a Song Model on Your Own Hardware. Read the License First. digitalmatters.me MiniMax Music 3.0 Puts a Song Model on Your Own Hardware. Read the License First. ojoo.ai's profile picture ojoo.ai 💽 Chinese MinMax has Released Music 3, An Open Source Music Generator It Creates Songs of up to 5 Minutes and Runs Locally - even on consumer hardware. [...] The Full Model fits on a GPU with 24 GB of VRAM, but it can run with as little as 8 GB. ➡ You can Try it for FREE on Hugging Face ➡ huggingface.co/space… 1 1 Log in to see more replies. Log in Log in or sign up for ThreadsSe