Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up

MiniMaxAI
/
MiniMax-Music3

Text-to-Audio
Diffusers
Safetensors
PyTorch
minimax_music3
music-generation
text-to-music
sglang-omni
Model card Files Files and versions
xet
Community
14

Instructions to use MiniMaxAI/MiniMax-Music3 with libraries, inference providers, notebooks, and local apps. Follow these links to get started.

  • Libraries
  • Diffusers

    How to use MiniMaxAI/MiniMax-Music3 with Diffusers:

    pip install -U diffusers transformers accelerate
    import torch
    from diffusers import DiffusionPipeline
    
    # switch to "mps" for apple devices
    pipe = DiffusionPipeline.from_pretrained("MiniMaxAI/MiniMax-Music3", dtype=torch.bfloat16, device_map="cuda")
    
    prompt = "Astronaut in a jungle, cold color palette, muted colors, detailed, 8k"
    image = pipe(prompt).images[0]
  • Notebooks
  • Google Colab
  • Kaggle
New discussion
Resources
  • PR & discussions documentation
  • Code of Conduct
  • Hub documentation

demo "η§‹εŽ»ζ˜₯ε‡ ε›ž" is both awesome and buggy

#15 opened 22 minutes ago by
J22

Apple Silicon MLX version proof of concept

#14 opened about 6 hours ago by
liminalsunset

How to get accent?

#13 opened about 8 hours ago by
ryg81

GGUF?

πŸ”₯ 1
1
#12 opened about 16 hours ago by
rodrigomt

Any plans to release a RVQ encoder or Flow-VAE encoder to enable audio-conditioned generation?

πŸ€— 7
1
#11 opened about 18 hours ago by
TheLatentSpacer

Is the model trainable?

πŸ‘ 1
6
#10 opened about 18 hours ago by
PabloFG

Without audio refrences this is useless, even more so for actual musicians

πŸ‘€ 1
1
#9 opened about 18 hours ago by
FreeDiddy

Imagine if MiniMax Speech 2.8 HD (and maybe Turbo) releases on Hugging Face!

#8 opened about 19 hours ago by
MihaiPopa-1

rδΈŠηœ‹εˆ°ηš„,η₯θ΄Ίε‘εΈƒ

#7 opened about 21 hours ago by
aifeifei798

Is the song reference supported?

3
#6 opened about 21 hours ago by
AlperKTS

RVQ encoder

πŸ‘πŸ‘€ 6
1
#5 opened about 21 hours ago by
arifqai

music quality look like suno V3.5

5
#4 opened about 21 hours ago by
rafiislam
Company
TOS Privacy About Careers
Website
Models Datasets Spaces Pricing Docs