Minimax Dropped A New Local Music Model
MiniMax released MiniMax-Music3 today (Aug 13, 2026), with open weights on Hugging Face, GitHub, and ModelScope.
The gist:
Give it a creative concept and optional lyrics, and it composes, arranges, performs, and produces a complete song in one generation, up to five minutes long. (MiniMax) Output is 32 kHz, 16-bit stereo. (ComfyUI Blog)
Architecture is a hybrid: an 8B "Global LLM" (initialized from Qwen3-8B) handles long-range structure, a 0.6B "Local LLM" fills in frame-level acoustic detail, and a 2.4B flow-matching synthesis stage fuses both models' hidden states instead of decoding from discrete tokens alone. (ComfyUI Blog)
Weights are under the MiniMax-Music3 Community License, and ComfyUI already has native support with an official text-to-music workflow (ComfyUI) — requires ComfyUI 0.33.0 or newer.
TL/DR: The dildo of justice always arrives unlubed