

Stable Diffusion
#18 in Text-zu-BildStability Ai · seit 2022-08-22 · 14× · zuletzt 15. Aug. 2026
Stable Diffusion ist eine Familie quelloffener Text-zu-Bild-Diffusionsmodelle von Stability AI, ursprünglich am 22. August 2022 veröffentlicht. Aktuell bildet die Modellreihe Stable Diffusion 3.5 (Large, Large Turbo, Medium) sowie die gehosteten Dienste "Stable Image Core" und "Stable Image Ultra" das Kernangebot. Die offenen Modelle können lokal (z.B. via Hugging Face/ComfyUI) unter der Stability AI Community License genutzt werden, während die gehosteten Varianten über eine kreditbasierte API (platform.stability.ai) sowie Cloud-Partner wie AWS Bedrock und Azure AI Foundry angeboten werden. Fine-Tuning ist über Methoden wie LoRA, DreamBooth und Textual Inversion sowie ControlNets möglich.
Features
| Fine-tuning | Unterstützt via LoRA, DreamBooth, Textual Inversion sowie ControlNets (z.B. Blur, Canny, Depth) für SD 3.5 |
| Generierungszeit | Ursprüngliches Stable Diffusion: Bilder in 512x512 in wenigen Sekunden auf Consumer-GPUs; Cloud-API-Dienste liefern typischerweise Ergebnisse in 8-12 Sekunden |
| Lizenz | Stability AI Community License: kostenlos für Forschung, nicht-kommerzielle Nutzung und Unternehmen mit unter 1 Mio. USD Jahresumsatz; darüber Enterprise License nötig |
| Max-Auflösung | Stable Diffusion 3.5 Large: 1 Megapixel; Medium: 0,25–2 Megapixel; Stable Image Core: 1,5 Megapixel; Stable Image Ultra: 1 Megapixel (Standard 1024x1024) |
| Plattform | Self-Hosting (Hugging Face, ComfyUI, lokale GPU) sowie gehostete API über platform.stability.ai, zusätzlich verfügbar via AWS Bedrock und Azure AI Foundry |
| Preis | API kreditbasiert: 1 Credit = $0,01; Stable Image Core $0,03/Bild, Stable Image Ultra $0,08/Bild; Open-Source-Modelle kostenlos selbst hostbar |
| Release-Datum | Ursprünglich 22. August 2022 (v1.4); aktuellste Modellfamilie Stable Diffusion 3.5 am 22. Oktober 2024 (Large/Large Turbo) bzw. 29. Oktober 2024 (Medium) |