Stable Diffusion locally: installation and models
Stable Diffusion is a family of open-weight image models, not an application: you run it in a local interface. To get started, use SDXL (6.94 GB file, 8 GB of comfortable VRAM) in Fooocus, ComfyUI Desktop, or, on Mac, Draw Things. SD 1.5 works for smaller cards; SD 3.5, announced on October 22, 2024, is for cards with 12 GB or more. Newer families, such as Flux or Z-Image, mainly run through ComfyUI.
Stability AI announced the public release of Stable Diffusion on August 22, 2022: a model that turns text into images on your own computer, without an account or volume limit. As of September 28, 2026, the name covers three generations of models, four or five possible interfaces, and licenses that differ substantially. This guide helps you choose a version and an interface, install everything on Windows, Mac, or Linux, and understand what you’re allowed to do with the images.
#What is Stable Diffusion?
Stable Diffusion is a latent diffusion model. It starts with an image of pure noise and denoises it step by step, guided by your text, until a coherent image emerges. The computation takes place in a compressed space called latent space, which explains why a consumer graphics card is sufficient where the first diffusion models required a computing center.
What made it famous wasn’t its raw quality but its openness. When it was released in 2022, Stability AI chose a CreativeML OpenRAIL-M license, which it describes as “permissive” and allowing both commercial and noncommercial use: the weights can be downloaded, modified, and retrained. Thousands of specialized variants exist. Stable Diffusion isn’t an application: it’s a model file that you run in the interface of your choice.
The name creates a vocabulary trap. The reference SD 1.5 repository on Hugging Face is now presented as a mirror of the old RunwayML repository, deprecated and unrelated to that company. And since 2024, the most discussed open image models have come from other labs: Flux, Qwen Image, or Z-Image. They run in the same interfaces, but they are not called Stable Diffusion.
#The versions, and which one to choose
AI images and videos on your own machine, no subscription and no credits: ComfyUI, Flux, Z-Image and Wan 2.2 with ready-to-load workflows, VRAM tiers, LoRA training and the legal frame.
- Lifetime online access
- PDF + files
- Lifetime updates
| Version | Output | Main file | Native resolution | Comfortable VRAM | Key takeaway |
|---|---|---|---|---|---|
| SD 1.5 | 2022 | 4.27 GB (v1-5-pruned-emaonly) | 512 × 512 | 4 to 6 GB | Lightweight, with a huge library of variants and LoRAs. Details and hands could be better. |
| SDXL 1.0 | July 2023 | 6.94 GB (sd_xl_base_1.0) | 1024 × 1024 | 8 GB: Stability AI claims it runs properly on 8 GB cards | The right starting point: significantly better quality, a mature ecosystem. |
| SD 3.5 | October 22, 2024 | Medium: 2.5 billion parameters; Large: 8.1 billion; Large Turbo: 4-step version | Medium: 0.25 to 2 megapixels; Large: 1 megapixel | Medium: 9.9 GB excluding text encoders (Stability AI). Large: more than 16 GB in 16-bit (our calculation) | Better prompt adherence. Different license; see below. |
Two details can help you avoid purchasing or downloading mistakes. First, SD 1.5 comes in two files: v1-5-pruned-emaonly.safetensors (4.27 GB), which the model card describes as using less VRAM and being suitable for inference, and v1-5-pruned.safetensors (7.7 GB), intended for retraining. For image generation, choose the first one. Next, SD 3.5 Large, with its 8.1 billion parameters, should not be evaluated based on the Medium model card: do not count on it running on an 8 GB card.
If you're starting with an 8 GB card, begin with SDXL: it offers the best balance of quality, speed, and available documentation. With 12 to 16 GB, also look at recent ComfyUI models: its README announces native support for Flux.1, Flux.2, Qwen Image, and Z-Image.
#The required hardware
| Your hardware | What works well |
|---|---|
| Card NVIDIA 4 to 6 GB (GTX 1660, RTX 2060, RTX 3050) | SD 1.5. Fooocus, under SDXL, lists 4 GB of VRAM and 8 GB of RAM as the minimum, relying on Windows virtual memory: slow, but possible. |
| NVIDIA card 8 GB (RTX 3060 Ti, RTX 4060, RTX 5060) | SDXL comfortably, threshold announced by Stability AI |
| NVIDIA card, 12 to 16 GB (RTX 3060 12 GB, RTX 4070, RTX 5070 Ti) | SDXL with LoRA and upscaling, SD 3.5 Medium, Z-Image Turbo (16 GB according to its listing), Flux in FP8 |
| Recent AMD card (RX 7000, RX 9000) | Good on Linux with ROCm; on Windows, official ROCm support in ComfyUI Desktop since January 2026 |
| Mac Apple Silicon | Any M-series chip: unified memory serves as VRAM. Slower than a NVIDIA card in the same performance class. |
| No graphics card | Technically possible, but much slower: reserve it for testing |
Plan for disk space too. A single SDXL file weighs 6,94 GB, and collections of LoRAs and variants grow quickly. Fooocus recommends having at least 40 GB free on each disk if you encounter the “RuntimeError: CPUAllocator” error.
#Which interface should you install?
| Interface | Who it's for | Strengths | Limitations |
|---|---|---|---|
| Fooocus | Beginners | A text field, a button, carefully chosen default settings; fewer than three clicks from download to the first image, according to its README | SDXL only: the project has limited long-term support, with bug fixes only |
| Forge (lllyasviel's repository) | Tabbed-interface users | The more memory-efficient AUTOMATIC1111 interface, with Flux support | No commits to the main branch since June 2025; development continues through forks such as Forge Neo |
| AUTOMATIC1111 | Those following older tutorials | The legacy interface, the best documented | Latest version released in February 2025, Flux not supported |
| ComfyUI | Advanced users and curious minds | Native support for recent models, reproducible workflows, very active project | Learning curve |
| Draw Things | Mac, iPhone, iPad | Native application, free edition that computes locally and offline, on-device LoRA training | Apple ecosystem only; paid cloud offerings are available as an option, but avoid them if you want to stay 100% local |
- Fooocus: installing it and generating your first image
- ComfyUI: installation and getting started
- AUTOMATIC1111: install Stable Diffusion WebUI
- Overview: generating images locally
#The shortest path per system
- 01Windows with a NVIDIA cardFastest: Fooocus. Download the archive from the official GitHub page, extract it, and run run.bat. On the first launch, the models are downloaded to Fooocus\models\checkpoints. For a tool that keeps up with new models, install ComfyUI Desktop instead, recommended for new users by the ComfyUI README.
- 02RTX 50 card: check PyTorchForge’s all-in-one archives list CUDA 12.1 and PyTorch 2.3.1 as the recommended versions. However, RTX 50 cards require a recent PyTorch version: an AUTOMATIC1111 contributor mentions PyTorch 2.7.0 as the first official version to support them. On an RTX 50, prefer ComfyUI, whose documentation requires CUDA 13.0 for recent NVIDIA cards.
- 03Mac Apple SiliconInstall Draw Things from the App Store: its free edition runs locally and offline, with no command line. Do not enable its cloud options if you want nothing to leave your Mac. ComfyUI Desktop is also available for macOS on Apple Silicon.
- 04Linux, or an AMD cardManual ComfyUI installation: clone the repository, set up a Python environment, and install the CUDA or ROCm version of PyTorch depending on the GPU. The detailed procedure is in our ComfyUI guide. Fooocus also offers a Linux installation via Anaconda.
- 05Verify that the GPU is workingDuring generation, nvidia-smi should show the VRAM in use and the GPU near 100%. If the image takes several minutes, the interface is running on the CPU: it is almost always an unsuitable PyTorch version.
#Where to find the models, and which ones to avoid
Two sources dominate. Hugging Face hosts official models from research labs. Civitai brings together community variants: fine-tuned models, LoRAs, and styles. A LoRA is a small file that adds a style or subject to a base model without replacing it. It must match the right family: an SDXL LoRA does not work on SD 1.5.
#Writing a prompt that works
- Describe; don't issue commands
- These models don't follow instructions. “Portrait of an older woman, window light, 85 mm lens, shallow depth of field” works; “take a beautiful photo of me” doesn't.
- Write in English
- The SD 1.5 model card says “Language(s): English” and specifies a frozen CLIP ViT-L/14 text encoder. French is understood, but less precisely.
- Subject, then framework, then style
- Order matters: what comes first carries the most weight.
- The negative prompt
- It lists what you don't want to see. It's especially useful for SD 1.5 and SDXL; check a recent model's card to see whether it takes this into account.
- Keep the seed
- Same seed, same settings, same image. To compare two prompts, fix the seed and change only one thing at a time.
#Licenses and commercial use
| Model | License | Commercial use of images |
|---|---|---|
| SD 1.5 | CreativeML OpenRAIL-M | Allowed: Stability AI describes it as permissive for commercial and noncommercial uses, with ethical-use restrictions listed in the license |
| SDXL 1.0 | CreativeML OpenRAIL++-M | Allowed, same principles |
| SD 3.5 | Stability AI Community License | Free for noncommercial use and for commercial use up to 1 million dollars in annual revenue; beyond that, contact Stability AI for an enterprise license. You retain ownership of the generated images. |
The model license does not settle everything. Copyright in a generated image and the use of a real person's likeness are governed by the laws of your country, not the software license. This guide describes the licenses published by their authors and does not constitute legal advice.
- The AI Image Kit: graphics-card settings and ready-to-use workflows
- Source: the announcement of Stable Diffusion 3.5 and its license
- Source: the announcement of the 2022 public release
#FAQ
Is Stable Diffusion free?+
Can Stable Diffusion be used without a graphics card?+
Which version should you choose: SD 1.5, SDXL, or SD 3.5?+
Do my images remain private?+
Stable Diffusion or Flux?+
Can you use Stable Diffusion on Mac?+
Can you sell generated images?+
Feedback, an error, or a clarification? Let us know—it improves the guide for everyone.