Comparison table
| SD 1.5 | SDXL 1.0 | SD 3 Medium | SD 3.5 Large | |
|---|---|---|---|---|
| Release | Oct 2022 | Jul 2023 | Jun 2024 | Oct 2024 |
| Parameters | 0.9 B | 3.5 B (+ refiner) | 2 B | 8 B |
| Native size | 512×512 | 1024×1024 | 1024×1024 | 1024×1024 |
| VRAM (fp16) | 4–6 GB | 8 GB | 10 GB | 16 GB+ |
| Speed (RTX 3060) | ~4 s | ~15 s | ~20 s | ~60 s |
| Prompt understanding | keywords | keywords, better composition | sentences, several subjects | best |
| Text in images | no | rarely | yes, short | yes |
| Artist names | strong | strong | weak | weak |
| Community models | thousands | many | few | growing |
| Licence | OpenRAIL-M | OpenRAIL++ | Community licence (free under $1M revenue) | |
| Try on this site | SD 1.5 | SDXL | SD 3 | — |
Stable Diffusion 1.5: the workhorse
Small, fast, endlessly customised. Photorealistic and anime fine-tunes, ControlNet models, LoRAs and embeddings exist for everything. Its weaknesses are the 512 px base resolution, hands, and a literal reading of prompts. It remains the right choice for a 6 GB card, for animation pipelines and for ControlNet-heavy work. This site's member txt2img runs SD 1.5.
SDXL 1.0: detail and style
Four times the pixels, a second text encoder, a refiner for the last steps. Faces, hands and composition improved a lot, and artist names have a strong, predictable effect, which is why our 1,731-artist comparison was generated with SDXL. It needs 8 GB and is three to four times slower than 1.5. Distilled versions (SDXL Turbo, LCM LoRA) bring it back to one second per image at some quality cost.
Stable Diffusion 3 and 3.5: understanding and text
A new architecture (diffusion transformer with three text encoders) that follows sentences instead of keyword lists: "a red cube on a blue sphere, to the left of a green cone" comes out right, and short text in quotes renders correctly. 3 Medium had anatomy problems that 3.5 fixed; 3.5 Large is the best open model of the family, 3.5 Medium the balanced one. Artist names and old negative prompt habits matter less. Try Stable Diffusion 3 Medium here.
Which one should you pick?
- Learning, 6 GB card, ControlNet, animation: SD 1.5.
- Illustration, portraits, artist styles, 8–12 GB: SDXL.
- Posters with text, complex scenes, 12–16 GB+: SD 3.5 Medium or Large.
- No GPU: the online tools on this site, no account needed.