

Stability Audio 3.0
#17 v AI generování hudbyStability Ai · v3.0 · 3× · naposledy 30. 6. 2026
Stability Audio 3.0 (officially "Stable Audio 3.0") is a model family for AI audio generation released by Stability AI in May 2026, consisting of four models: Small SFX, Small, Medium, and Large. The models generate instrumental music, sound design, and sound effects from text prompts, with variable length from a few seconds to over six minutes for Medium and Large. Three of the four models (Small SFX, Small, Medium) are available as Open Source weights on Hugging Face, while Large is only accessible via the Stability AI API or Enterprise Self-Hosting. The model does not generate intelligible singing voices or lyrics and offers no voice cloning; the training data comes entirely from licensed sources (including AudioSparx, Freesound).
Vlastnosti
| Echtzeit-Streaming | Nicht als Feature dokumentiert; Small-Modelle laufen on-device (Smartphones/Laptops) für lokale Generierung ohne Cloud-Anbindung |
| Latenz | Large-Modell für 'low-latency generation at high volume' auf Musikplattformen konzipiert; keine konkrete Zeitangabe veröffentlicht |
| Lizenz | Stability AI Community License (freie Kommerzialisierung der Outputs); Enterprise License Pflicht für Organisationen mit über $1 Mio. Jahresumsatz |
| Plattform | Hugging Face (Small SFX, Small, Medium als offene Gewichte), Stability AI API und Enterprise-Self-Hosting (Large), Web-App stableaudio.com, ComfyUI-Integration |
| Preis | Keine offizielle Preisliste für Stable Audio 3.0 veröffentlicht; API nutzt Credit-System (1 Credit = $0,01), Web-App bietet kostenlose Testcredits |
| Release-Datum | 20. Mai 2026 |
| Voice-Cloning | Nicht unterstützt – Modellfamilie ist auf Instrumentalmusik und Soundeffekte ausgelegt, keine verständlichen Gesangsstimmen oder Voice-Cloning |