音乐创作者和 AI 音频开发者终于有了一个可商用、可定制的开源音频模型——Stable Audio 3.0 支持六分钟生成和 LoRa 微调,做音乐生成或声音设计的团队可以直接上手实验。
Stability AI 推出了 Stable Audio 3.0,这是一个开源权重模型系列,专为艺术实验设计。新版本支持最长六分钟的变长音频生成,并能在便携设备上完成完整歌曲创作,无需 GPU。模型基于完全许可的数据集训练,用户可商用输出,年收入不超过 100 万美元。首次支持 LoRa 训练,允许用户用自己的音频库定制模型。Stability AI 邀请开发者参与实验,认为最佳创新仍在等待被构建。
Meet Stable Audio 3.0, the open-weight model famil…
Meet Stable Audio 3.0, the open-weight model family built for artistic experimentation.
This is our open invitation to experiment with generative audio. We believe the best innovations are still waiting to be built.
The 4-1-1 on 3.0:
📣 You own your outputs, and can distribute and commercialize them under the Stability AI Community License (up to $1 million in revenue).
🎵 New and improved capabilities include variable-length generation up to six minutes, and full song composition on portable devices, no GPU required.
✅ Trained on a fully licensed dataset.
🎨 You can customize the models on your own library with support for LoRa training, which we’ve documented for the first time.
More on the models 👇