heartmula

HeartMuLa

Most open-source music models give you one capability. HeartMuLa gives you four: a lyrics-conditioned song generator, a high-fidelity music codec, a lyrics transcription model, and an audio-text alignment model — all open-sourced together as a coherent foundation. The 3B generator handles multilingual lyrics across English, Chinese, Japanese, Korean, and Spanish, with style controlled through simple comma-separated tags. An internal 7B version already reaches Suno-level quality, with the open 7B release planned.

Apache-2.0
Text-to-Audio
PyTorch
Safetensors
Diffusers
by @AIOZAI
41
0

Last updated: 3 months ago


Sign in to see model files

orCreate an account